New essay: AI scales worse with more tokens than humans do with more time, but is catching up

tobyordoxford · x · 2026-10-09

Toby Ord (Oxford) published a new essay asking whether AI models scale as well with more tokens as humans do with more time. Reviewing the limited available evidence — including a METR chart plotting human scaling with time against AI scaling with tokens — he concludes that AI has so far scaled less well with more tokens than humans do with more time, but appears to be catching up. The claim has direct implications for forecasting how long-task agents improve as test-time compute grows.

Related event: Oxford's Toby Ord: AI Token Scaling Still Lags Human Time Scaling(5 posts)→

Original post →

More from AGI Musings

AGI Musings channel →