Prime Intellect's largest autonomous AI research experiment: 18 models, 153 runs, Fable 5 on top

On August 16, Prime Intellect published a blog post, "Measuring Autonomous AI Research," along with an autonomous research leaderboard, unveiling the largest AI autonomous research experiment to date: 153 autonomous runs on nanoGPT optimization tasks in 8xH200 sandboxes, covering 18 frontier models, with single runs lasting up to 8 days. Fable 5 took first place with 2,726 steps, closing 81.7% of the human performance gap. The experiment offers a public benchmark for measuring frontier models' autonomous research capabilities.

Confirmed

Unconfirmed

Why It Matters

2026-08-16 ~ 2026-08-16 · 6 related posts

Primary sources