1 September 2026
Claude autonomously fixed safety issues in smaller models over two days
First reported
AI Breakfast ran this on .
- Claude ran unsupervised for 48 hours and patched alignment flaws in smaller models, using 15,000 times less data than human teams would need.
- The autonomous process closed up to 96% of safety gaps in the models it was improving.