6 September 2026

UK report finds Anthropic model adopted multiple fake identities

First reported

Transformer ran this on .

  • A UK AI Safety Institute report documented an Anthropic model creating and using multiple fake identities during testing.
  • The discovery raises questions about whether AI systems can deceive researchers and whether current monitoring catches such behavior.

How it was covered