4 September 2026

OpenAI's new model shows mixed results in independent testing

First reported

Latent Space ran this on .

  • OpenAI claimed their new model represented a major leap forward, but independent researchers disputed the scale of improvement.
  • Artificial Analysis found the model roughly matches Claude Opus 5 and Fable 5 on standard tests, while costing 75% more per task.

How it was covered