Halved slope may be an artifact of higher-difficulty-variance dataset, log scaling still holds
tobyordoxford · x · 2026-09-14
tobyordoxford offers a best guess for the halved inference-scaling slope: the new dataset's problems have more variance in difficulty, which could easily explain the slope change — though not the large between-generation jump. He notes logarithmic inference scaling still holds either way.
Related event: Oxford's Toby Ord: 10x Compute in RLVR Yields ~3x Token Efficiency(3 posts)→
More from Models
- Leaked Astra and Fable sizes are far below 10T, says X user: scale barely maps to capability now — teortaxesTex · 2026-09-14
- Free ChatGPT solves a decade-old maths dice problem in 13 minutes — Chris_Armstrong · 2026-09-14
- David Bellamy clarifies his experiment used K2 Horizon, an open-weights 375B LLM — JeremyNguyenPhD · 2026-09-14
- Reddit user gets ChatGPT 'Cyber Abuse' warning, appeal denied in an hour with no explanation — TheBeaconCrafter · 2026-09-14
- Swift-Qwen3.8-27b, a token-efficient reasoning Qwen finetune, trends on Hugging Face — ukisai · 2026-09-14
- GPT-6 Astra hands-on: composes first, orchestrates later, and reportedly outshines Fable and Sol — paw_lean · 2026-09-14