Researchers flag massive reporting bias in AI math capabilities: failures go untracked

RexDouglass · x · 2026-09-08

Wes Pegden points out massive reporting bias in AI-in-math capabilities: nearly every time an agent solves a hard problem humans hadn't cracked gets reported, while failures in the other direction are never tracked — "marketing, not research." Rex Douglass adds that math has effectively zero metascience tracking its human failure rate, leaving the field with no baseline to judge AI claims against.

Original post →

More from Research

Research channel →