The Story of Inference-Time Scaling
(2026)…and the training-time loop behind it. A language model can get dramatically better answers by thinking longer at answer time — not only by being retrained or made larger.
Shashank Bangalore Lakshman, Scott Eiers. The Story of Inference-Time Scaling. 9th Workshop on Visualization for AI Explainability (VISxAI), IEEE VIS 2026. Accepted; to be presented.