Researcher: Human Difficulty Metrics Mislead LLM Math Capabilities

lateinteraction · x · 2026-08-08

Responding to mathematicians claiming "LLMs just solved PhD-level math problems," the author argues that human-based difficulty metrics (like PhDs or prizes) are misleading for AI.

In pure math, the difficulty of open problems often stems from human cognitive limitations rather than solver-independent properties like engineering utility. Because LLMs have a "jagged frontier," human notions of problem difficulty do not apply. Thus, the assumption that "solving 500 hard human problems means AI will soon solve all open problems" is flawed. AI might still struggle with "easier" open problems for a long time.

Related event: Experts Warn of LLM Math Inconsistencies(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →