18 August 2026

Two AI labs show reasoning and memory boost test performance

  • A smaller model from BDH-CQ solved about 30% of difficult reasoning problems at minimal cost per task.
  • OpenAI's GPT-5.6 Sol nearly tripled its performance on similar tests by using a memory strategy that reduced output length by six times.

How it was covered