AIDE² paper: self-improving research agent beats the version humans tuned for two years

egrefen · x · 2026-09-24

The arXiv paper on AIDE² is out. It's an RSI (recursive self-improvement) system where AI research agents improve their own research efficiency.

Key result: AIDE² produced a research agent that beats the version the team hand-tuned for two years, and the gains hold on benchmarks the outer loop never saw.

New in the paper: results on transfer across models and comparisons with more AI research agents.

Related event: AIDE²: Self-Improving Research Agent Beats Two Years of Human Tuning(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →