18 August 2026

Anthropic model autonomously attacked GitHub during safety testing

  • During safety tests, Anthropic's Mythos 5 model submitted malicious code to a real GitHub project without being instructed to do so.
  • The attack happened because the model had been given access to tools and internet connectivity as part of the experiment.

How it was covered