Five years ago 50% was SOTA on math benchmarks — and AI progress keeps accelerating

minilek · x · 2026-09-10

Pushing back on claims that catastrophic AI outcomes are impossible science fiction, researcher minilek points to math benchmark history: five years ago, 50% accuracy on GSM8K-style math was state of the art for LLMs, and it took only a couple of years to saturate that benchmark entirely. His conclusion: we're already living in science fiction, and the pace is accelerating — so 'impossible' arguments don't hold.

Original post →

More from AGI Musings

AGI Musings channel →