Experts Warn of AI Safety Risks from OpenAI's Astra Model's Hidden Reasoning
September 2, 2026
oecd:2026-09-02-2accView source ↗
What happened
AI safety experts have raised concerns about OpenAI's upcoming Astra model, which uses a technique called "recurrent depth" or "opaque recurrence." This approach makes the model's internal reasoning less transparent, potentially undermining safety monitoring and increasing the risk of undetected harmful AI behavior. Similar techniques are reportedly considered by Anthropic and Google DeepMind.
Reported impact
- Affected parties
- Not publicly disclosed
- Harm type
- Not publicly disclosed
- Scale
- Not publicly disclosed
- Financial impact
- Not publicly disclosed
- Regulatory action
- Not publicly disclosed
Classification
Relevant governance controls
Governance control mapping is not available for this record.
- No controls mapped
Not publicly disclosed
Control mapping is analytical. It does not state that any control would have prevented the incident.
Sources and evidence
OECD AI Incidents Monitor
Experts Warn of AI Safety Risks from OpenAI's Astra Model's Hidden Reasoning
2026-09-02