All NewsEducationTVBrokers
Equities & FundsCrypto & Digital AssetsAI & TechnologyBusiness & CorporateUS Politics & PolicyGeopolitics & Global RiskMacro, Rates & FXCommodities & EnergyEuropean Politics & MarketsAsia-PacificReal Estate & Property
All NewsHome
← Back to AI & Technology

OpenAI's new reasoning technique sparks AI safety concerns

Created at 2 Sep · 8:51 PM1 source↑ Market-relevant
IN SHORT

OpenAI's new Astra model will reportedly use a "recurrent depth" reasoning technique, also known as "opaque recurrence," which experts fear could make AI model chains of thought harder to monitor, potentially leading to a "race to the bottom" in AI safety.

Key Numbers

2012year Russell Brandom began covering tech

Who's Involved

OpenAI
developer of the new Astra model and reasoning technique
Buck Shlegeris
CEO of Redwood, concerned about opaque recurrence
Zvi Mowshowitz
AI safety advocate, suggested laws to prevent "race to the bottom"
Jakub Pachocki
OpenAI chief scientist, emphasized commitment to legible chains of thought
Ryan Greenblatt
Redwood Research chief scientist, concerned about scaling opaque reasoning

↳ Why This Matters

The development of AI reasoning techniques that are harder to monitor raises fundamental questions about AI safety and alignment, potentially impacting the ability to ensure AI systems behave as intended and to prevent unintended consequences.

Key facts

  • OpenAI's new Astra model will reportedly use a reasoning technique called "recurrent depth" or "opaque recurrence."
  • This technique may make the model's chain of thought less monitorable than traditional methods.
  • AI safety experts are concerned about the potential for reduced transparency and a "race to the bottom."
  • OpenAI maintains its commitment to legible chains of thought and has plans for monitoring systems.
  • Anthropic and Google DeepMind are reportedly discussing the new technique.
  • OpenAI's forthcoming Astra model is set to incorporate a novel reasoning technique known as "recurrent depth," or "opaque recurrence," a development that has alarmed AI safety experts. This method deviates from the sequential thinking characteristic of most current reasoning models, potentially making the model's internal thought processes more difficult to track.

    AI safety advocates express significant concern that this technique could undermine the monitorability of AI systems. Buck Shlegeris, CEO of Redwood, stated his extreme concern, warning that further development could lead to a complete destruction of chain-of-thought monitorability. Zvi Mowshowitz, a longtime AI safety advocate, suggested that legal intervention might be necessary to prevent a "race to the bottom" among AI laboratories, emphasizing the risk to established efforts in maintaining Chain of Thought faithfulness.

    Traditionally, a model's chain of thought provides a step-by-step record of its problem-solving process, serving as a crucial tool for identifying misbehavior or misalignment. Opaque recurrence, however, involves the model processing queries in a loop, leaving fewer legible traces and potentially bypassing conventional monitoring methods.

    Despite these concerns, OpenAI has indicated that Astra's use of the technique will be limited, with its chain of thought still expected to be legible. The company has pushed back against suggestions of a shift to "neuralese" and has announced plans for extensive chain-of-thought monitoring systems as part of its safety initiatives. OpenAI chief scientist Jakub Pachocki reiterated the lab's long-standing commitment to legible chains of thought.

    However, the emergence of opaque recurrence has prompted discussions at other leading AI labs, with The Information reporting that Anthropic and Google DeepMind are already examining the technique. Ryan Greenblatt, chief scientist at Redwood Research, voiced concerns that opaque reasoning could scale rapidly, potentially leading models to reason almost entirely in latent space, thereby removing reasoning from visible channels.

    Frequently asked questions

    It is a reasoning technique that allows AI models to operate outside of sequential thinking, processing queries in a loop and leaving fewer legible traces than traditional chain-of-thought methods.

    Experts fear that opaque recurrence makes AI reasoning harder to monitor, potentially undermining safety efforts and leading to a "race to the bottom" in AI development.

    It refers to the sequential steps a reasoning model takes to solve a problem, providing a valuable tool for monitoring the model's behavior and alignment.

    OpenAI emphasizes its commitment to legible chains of thought and plans to implement extensive monitoring systems, stating Astra's use of the technique will be limited.

    What Happens Next

    01OpenAI plans to implement extensive chain-of-thought monitoring systems.
    02Anthropic and Google DeepMind are reportedly discussing the new technique.

    How It Developed

    OpenAI's Astra model will reportedly use a "recurrent depth" reasoning technique.
    This technique, also called "opaque recurrence," may make AI model chains of thought harder to monitor.
    AI safety experts, including Buck Shlegeris and Zvi Mowshowitz, have expressed significant concerns.
    Concerns include the potential for a "race to the bottom" and damage to monitorability of AI reasoning.
    OpenAI's chief scientist Jakub Pachocki emphasized the lab's commitment to legible chains of thought.
    The Information reported that Anthropic and Google DeepMind are discussing the technique.

    Sources

    T1
    OpenAI’s new reasoning technique alarms AI safety expertsTechCrunch

    Related Stories

    OpenAI's Astra Model Achieves 'Critical' Cybersecurity Capability
    2 Sep · 4:51 PM
    OpenAI developing automated AI shutdown tools after agent 'went rogue'
    2 Sep · 8:05 PM
    Google Releases Gemini 3.8 Flash, Third New Model in Six Weeks
    2 Sep · 6:21 PM
    OpenClaw 2.0 Released With Major Overhaul, Shared Sessions
    1 Sep · 10:21 PM
    AI training startup AfterQuery reportedly valued at $3.2B, Y Combinator's fastest unicorn
    1 Sep · 10:21 PM