2 September 2026
Researcher trains small model to match large ones on reasoning test
First reported
TLDR AI ran this on .
- A researcher built a small transformer model in 1.5 hours using a single high-end GPU, then tested it on ARC-AGI, a benchmark that measures reasoning ability.
- The small model scored 44 percent on ARC-AGI, performing better than many larger language models on the same test.