AIDE² paper shows AI research agent recursively rewriting its own code, 7 gains in 8-day run
krishnan · x · 2026-10-02
A new paper, "Recursive self-improvement of AI research agents," demonstrates genuinely recursive self-improvement.
How it works:
- Instead of a fixed program editing others' code, AIDE² treats its own source code as the thing being edited
- It proposes changes to itself, builds that version, runs it on AI R&D tasks, and keeps whatever scores best on a hidden evaluation
- Each accepted rewrite becomes the agent performing the next round of editing — hence "recursive"
Results:
- In one 8-day autonomous run it found seven successive improvements, from a new search policy to memory mechanisms that compress and manage its own growing context
- Gains held on four held-out benchmarks spanning ML engineering, heuristic algorithm engineering, and physics-based weather forecasting
Related event: AI Research Agent Achieves Recursive Self-Improvement(2 posts)→
More from coding & agent
- Jev pitches 'decision primitives': models plugging into logic without text — hardimanjames · 2026-10-02
- LangChain's Sproul: agent core pattern unchanged for a year, "we've been at AGI for four months" — BraceSproul · 2026-10-02
- n8n integrates typesafeai's Jev model as a smarter If/Switch for workflows — hardimanjames · 2026-10-02
- Firecrawl launches People Enrichment Pack so agents can find buyers, talent and company data — devdigest · 2026-10-02
- Stanford launches CS 224V, an Agentic AI course tackling agent reliability with RAG and formal methods — stanfordnlp · 2026-10-02
- Early Hands-On: OpenAI's Dots Agent Impresses With Speed and First-Try Accuracy — billyjhowell · 2026-10-02