2 September 2026

Researcher trains small model to match large ones on reasoning test

First reported

TLDR AI ran this on .

  • A researcher built a small transformer model in 1.5 hours using a single high-end GPU, then tested it on ARC-AGI, a benchmark that measures reasoning ability.
  • The small model scored 44 percent on ARC-AGI, performing better than many larger language models on the same test.

How it was covered