/Submit incident
Documented

Experts Warn of AI Safety Risks from OpenAI's Astra Model's Hidden Reasoning

September 2, 2026
oecd:2026-09-02-2accView source ↗

What happened

AI safety experts have raised concerns about OpenAI's upcoming Astra model, which uses a technique called "recurrent depth" or "opaque recurrence." This approach makes the model's internal reasoning less transparent, potentially undermining safety monitoring and increasing the risk of undetected harmful AI behavior. Similar techniques are reportedly considered by Anthropic and Google DeepMind.

Reported impact

Affected parties
Not publicly disclosed
Harm type
Not publicly disclosed
Scale
Not publicly disclosed
Financial impact
Not publicly disclosed
Regulatory action
Not publicly disclosed

Classification

Organization
Not publicly disclosed
AI system
Not publicly disclosed
Industry
Not publicly disclosed
Country
Not publicly disclosed
Provider
Not publicly disclosed
Incident type
Not publicly disclosed

Relevant governance controls

Governance control mapping is not available for this record.

  • No controls mappedNot publicly disclosed

Control mapping is analytical. It does not state that any control would have prevented the incident.

Sources and evidence

OECD AI Incidents Monitor
Primary source
Experts Warn of AI Safety Risks from OpenAI's Astra Model's Hidden Reasoning
2026-09-02