2 September 2026

OpenAI's Astra model reaches highest cybersecurity risk category

First reported

MIT Technology Review ran this on , a day before the other 8 sources picked it up.

  • Astra became OpenAI's first model to reach the 'Critical' threshold in its risk framework, meaning it can find and exploit previously unknown security flaws without human guidance.
  • The model uses a technique called recurrent depth where it analyzes text in repeated loops to improve reasoning, but this makes its decision-making harder for humans to monitor.

Where they differ

Reported by CNBC, OpenAI