François Chollet Clarifies LLM Stance: Test-Time Compute Breaks Base Model Bottlenecks

fchollet · x · 2026-08-02

François Chollet clarified his evolving stance on Large Language Models (LLMs), stating that his past criticisms of base LLMs do not apply to systems utilizing test-time compute (TTA). He compared the difference to steam trains versus modern electric bullet trains—similar in appearance, but fundamentally different in underlying principles.

Chollet noted that he updated his views in December 2024, no longer believing that the LLM tech platform (enhanced with TTA) would stall. He emphasized that if we look strictly at base models without TTA, they still perform poorly on the 2019 ARC 1 benchmark, despite roughly a 100,000x increase in compute scaling. The paradigm shift to test-time compute, rather than simply scaling single-pass static inference, is what enables today's state-of-the-art systems to achieve advanced reasoning.

Related event: François Chollet: Test-Time Compute is Key as Pure Parameter Scaling Hits Limits(4 posts)→

Original post →

More from AGI Musings

AGI Musings channel →