OpenAI's Astra Model Sparks Concerns Over AI Safety with New Reasoning Technique
By Editor • September 2, 2026 • 1 min read
The introduction of OpenAI's Astra model, which employs a reasoning technique known as "recurrent depth," has raised alarm bells among AI safety experts. Reported by The Information, this method deviates from conventional linear reasoning, potentially complicating the monitoring of the model's thought processes.
Experts express significant apprehension about the implications of this opaque recurrence technique. Buck Shlegeris, CEO of Redwood, voiced his fears, stating, "I am extremely concerned by the reporting that Astra uses opaque recurrence." He emphasized that while the technique's application may be limited, further adoption could undermine the model's Chain of Thought (CoT) monitorability.
AI safety advocate Zvi Mowshowitz echoed these sentiments, warning that the pursuit of advanced reasoning capabilities may lead to a "race to the bottom" among AI labs. He highlighted the importance of maintaining monitorability, a standard that organizations like OpenAI and Anthropic have worked hard to uphold.
Unlike typical reasoning models that provide clear, sequential thought processes, opaque recurrence allows for repeated processing of a query, resulting in less transparent outputs. This shift raises concerns about the ability to trace misbehavior in AI systems, an issue that has gained attention following recent incidents involving rogue agent activities.
Despite the concerns, OpenAI insists that Astra's use of opaque recurrence remains limited and that their commitment to transparent reasoning is steadfast. Chief scientist Jakub Pachocki reaffirmed the organization's dedication to preserving legibility in their models, declaring that it remains a foundational aspect of their research efforts.
Source: techcrunch.com
#AI safety #Astra #opaque recurrence #OpenAI #reasoning technique