If this is true, the hyperscalers are toast
Summary
The author argues that small language models running locally may supplant cloud-based LLMs, citing Stanford research showing SLMs matching LLMs in many tasks at lower energy costs. The piece outlines potential investment implications for hyperscalers, GPUs, and data centers, while noting limits in agentic AI and on-device capabilities.