2 September 2026
OpenAI's Astra model reaches highest cybersecurity risk category
First reported
MIT Technology Review ran this on , a day before the other 8 sources picked it up.
- Astra became OpenAI's first model to reach the 'Critical' threshold in its risk framework, meaning it can find and exploit previously unknown security flaws without human guidance.
- The model uses a technique called recurrent depth where it analyzes text in repeated loops to improve reasoning, but this makes its decision-making harder for humans to monitor.