Two years from o1-preview to superhuman math: RL scaling now cracks open research problems

jam3scampbell · x · 2026-09-09

jam3scampbell notes it took just two years to go from o1-preview's AIME test-time scaling (Sept 2024) to superhuman performance on an evaluation set of open math problems. He argues GPT-6 Astra's 'internal model' already shows a step-change just days after release, and that honest extrapolation points to fully automated research in the not-too-distant future — urging people to internalize this rate of progress and act with urgency.

Original post →

More from AGI Musings

AGI Musings channel →