Largest open experiment: frontier models close 82% gap to human record in autonomous AI research

inductionheads · x · 2026-08-17

PrimeIntellect ran the largest open experiment on how frontier models do AI research. 100+ autonomous runs across 10+ models, sandboxed on 8xH200s for up to 8 days, iterating on the nanoGPT optimizer track. Best runs closed 82% of the gap to a record built by dozens of humans over months. FakePsyho comments that optimization problems are better at testing creativity and thus better proxies for RSI than ML problems.

Related event: Results of Largest Autonomous AI Research Experiment Open Sourced(13 posts)→

Original post →

More from AGI Musings

AGI Musings channel →