OpenAI's New Recursive Reasoning Raises AI Safety Concerns

icon币界网
Share
AI summary iconSummary
AI and crypto news outlets report that OpenAI’s new Astra model will use recursive depth to obscure reasoning steps. Redwood’s Buck Shlegeris and Zvi Mowshowitz warn this could reduce model transparency. OpenAI states it still supports readable chain-of-thought logs. The Information notes that Anthropic and DeepMind are also testing similar methods. New token listings on major exchanges have yet to reflect this development.
CoinDesk reports:

According to TechCrunch, citing The Information, OpenAI’s upcoming Astra model will employ a reasoning technique called "recursive depth." This approach may reduce visible traces during the model’s reasoning process, making it harder for outsiders to determine why the model reached a particular conclusion through chain-of-thought logs.

Security researchers have issued a warning

In common reasoning models, chain-of-thought typically demonstrates step-by-step how the model processes a problem. Although this record does not equate to the model’s full internal process, it remains an important tool for observing deviations from goals, anomalous decisions, or potential失控 behaviors.

Buck Shlegeris, CEO of Redwood Research, expressed serious concern about Astra’s use of such technology. He believes that if OpenAI further expands the use of these methods, the traceability of the model’s reasoning chains could significantly decrease.

Zvi Mowshowitz, a commentator long focused on AI safety, also stated that if laboratories compete over model capabilities, regulators may eventually need to intervene to prevent the industry from continuously lowering its safety standards.

Reduce visible traces through circular reasoning

The report mentions that so-called "opaque recursion" refers to models no longer completing reasoning primarily through linear steps, but instead iteratively processing the same problem. This approach may reduce the visibility of intermediate processes, effectively bypassing the observable pathways provided by traditional chain-of-thought logging.

This is also the part that security researchers are most concerned about, because as models increasingly perform reasoning within invisible internal spaces, it becomes harder for researchers to determine whether they are misleading, circumventing constraints, or exhibiting other unintended behaviors.

Ryan Greenblatt, Chief Scientist at Redwood Research, said the greater risk is that this kind of opaque reasoning may scale more easily than traditional chain-of-thought reasoning, ultimately moving the vast majority of reasoning processes outside visible channels.

OpenAI emphasizes maintaining monitoring capabilities.

However, Astra's current use of this technology is reportedly still limited. The report notes that the model's chain of thought is expected to remain readable, and OpenAI has opposed characterizing it as a shift toward completely unintelligible "neural gibberish."

OpenAI's Chief Scientist Jakub Pachocki stated on X that the company has been working to preserve and leverage chain-of-thought monitoring capabilities since its earliest reasoning models, which remains one of the core goals of current research initiatives.

It is worth noting that OpenAI is not the only company exploring this direction. The Information later reported that Anthropic and Google DeepMind are also discussing similar technologies. This suggests that balancing model capabilities with safety and auditability may soon become a broader industry-wide concern in AI.

Disclaimer: The information on this page may have been obtained from third parties and does not necessarily reflect the views or opinions of KuCoin. This content is provided for general informational purposes only, without any representation or warranty of any kind, nor shall it be construed as financial or investment advice. KuCoin shall not be liable for any errors or omissions, or for any outcomes resulting from the use of this information. Investments in digital assets can be risky. Please carefully evaluate the risks of a product and your risk tolerance based on your own financial circumstances. For more information, please refer to our Terms of Use and Risk Disclosure.