4 September 2026
OpenAI's new model shows mixed results in independent testing
First reported
Latent Space ran this on .
- OpenAI claimed their new model represented a major leap forward, but independent researchers disputed the scale of improvement.
- Artificial Analysis found the model roughly matches Claude Opus 5 and Fable 5 on standard tests, while costing 75% more per task.