4 September 2026

Apple researchers find language models skip ideal probability updates

First reported

Deep Learning Weekly ran this on .

  • Apple Machine Learning Research discovered that large language models, when given new information, do not follow optimal probability theory (Bayesian updates) the way mathematicians would expect.
  • Despite this deviation, the models' actual approach often produces better results on practical tasks than following the mathematically perfect method would.

How it was covered