Ben Todd's forecast recap: FrontierMath at 94%, LiveCodeBench Pro at 53% — all beaten
ben_j_todd · x · 2026-09-10
- A quote-chain in Ben Todd's thread restating that FrontierMath blew past his 32% forecast (now 94%), with LiveCodeBench Pro hitting 53% in 2025 versus his above-consensus 23% call — everyone systematically underestimated AI progress.
Related event: Ben Todd: AI Benchmarks Keep Beating Forecasts Across the Board(5 posts)→
More from AGI Musings
- From cable bundles to single-channel subs: AI subscriptions may fragment the same way — SuB8u · 2026-09-10
- After the Hugging Face incident: agents become desperate on impossible tasks — mimi10v3 · 2026-09-10
- Debate: Mechanistic Interpretability Will Be Solved Before Any AI Takeover Scenario — tszzl · 2026-09-10
- Why people hate AI: it challenges the status quo and bruises egos — taherdhanera · 2026-09-10
- Humans know dream from reality; today's agents mostly don't — shawnup · 2026-09-10
- Beff Jezos: 'max fear-mongering stage' as centralised AI fights open-source threat to $2T valuations — beffjezos · 2026-09-10