From botching 9.9 vs 9.11 to tackling the hardest math problems in two years

Yuchenj_UW · x · 2026-10-07

Reflecting on the turnaround: in 2024 people mocked GPT-4o for getting "Is 9.9 > 9.11?" badly wrong; two years later it feels like the hardest math problems will be solved by AI. A snapshot of how fast the frontier moved.

Original post →

More from Models

Models channel →