Models Are Getting Dumber on Purpose
Summary
The article argues that large language models improve reasoning on benchmarks but struggle with factual recall, and it advocates using retrieval, tools, and grounding to mitigate hallucinations. It also discusses frontier models that can run on consumer GPUs with a small active parameter set, enabling on-device operation and new ways to manage up-to-date knowledge.