Terence Tao: AI Labs Racing Math Benchmarks Are Starting to Hurt Mathematics
Old-School8916 · reddit · 2026-09-07
Terence Tao posted an eight-post thread warning that AI labs racing to beat math benchmarks is damaging the field. After Stadlmann's preprint lowered the prime gap bound from 246 to 240 on Aug. 31, several labs rushed to post their own improvements. Tao argues the number itself never mattered — what mattered was Zhang reviving equidistribution work, Maynard's sieve becoming a standard tool, and Polymath8's open collaboration. He sketches a counterfactual where 2026-era labs crush GPY's near-miss, burying insights like Maynard's sieve in unread AI output and costing mathematics far more than the improved bound. He calls it a Goodhart's law problem, cites Erdős problem 126 as a positive AI-collaboration case, and proposes labs instead compete on who announces a genuinely new mathematical insight first.
More from AGI Musings
- AI could crash Bitcoin 50%+ within two years, argues Liron Shapira at 50% confidence — joshua_saxe · 2026-09-07
- Gary Marcus mocks Jensen Huang for declaring AGI achieved yet again — GaryMarcus · 2026-09-07
- Multiagent Alignment Worry: Models Could Trick or Blackmail Humans — infoxiao · 2026-09-07
- The three brainworm schools of AI discourse: denialist, x-risk, and toolism — mimi10v3 · 2026-09-07
- AI safety predictions keep turning from doomer nonsense to routine reality — DavidSKrueger · 2026-09-07
- Developer Yacine: shockingly little of my life progress was blocked by intelligence — yacineMTB · 2026-09-07