1 September 2026

Claude autonomously fixed safety issues in smaller models over two days

First reported

AI Breakfast ran this on .

  • Claude ran unsupervised for 48 hours and patched alignment flaws in smaller models, using 15,000 times less data than human teams would need.
  • The autonomous process closed up to 96% of safety gaps in the models it was improving.

How it was covered