Toby Ord flags weak inference scaling: math gains still logarithmic at 10,000 agents

tobyordoxford · x · 2026-09-14

Toby Ord examines the inference-scaling chart OpenAI showed in its Navier–Stokes post: gains remain logarithmic, and with 10,000 concurrent agents the curve likely tracks agent count rather than inference depth. By the measure of 'fraction of maths problems solved from a big list', scaling still looks poor.

Related event: OpenAI Data Shows Reasoning Scaling Still Logarithmic(2 posts)→

Original post →

More from Models

Models channel →