Experimental Evidence of RSI: Self-Improving Research Agents
JensHonack · x · 2026-07-16
A cited post discusses experimental results regarding Recursive Self-Improvement (RSI): the author claims that over 8 consecutive days, they allowed an "automated research agent for automated research" to continuously optimize itself. It ultimately outperformed a manually tuned harness on a holdout set benchmark, building upon two years of manual parameter tuning.
The core focus isn't making broad AGI claims, but rather placing "self-improvement" within a measurable experimental framework, highlighting that:
- Automated optimization processes can already outperform manually engineered versions on specific tasks.
- The results are based on comparisons against holdout set benchmarks.
- This kind of work will spark further discussion on the feasibility of RSI.
Related event: AIDE² Self-Improvement Run Beats 2 Years of Manual Tuning(12 posts)→
More from AGI Musings
- Accelerationist fires back at AI doomers: beliefs aren't arguments — Dan_Jeffries1 · 2026-09-11
- "ChatGPT 6 Makes Workers with IQ Below 130 Useless": French AI Debate Sparks Backlash — mitchdeg · 2026-09-11
- 'AGI is here' vs reality: AI labs still ship some of the jankiest desktop apps ever — MilesCranmer · 2026-09-11
- Harry Collins: LLMs can't do frontier science because they can't invent new language — whoamisri · 2026-09-11
- The Waymo effect: how AI is quietly making research less collaborative — JohnHammersley · 2026-09-11
- Misquoted: Anthropic Staff Warned of Double-Digit Extinction Risk by 2030, Not Dismissed It — davidmanheim · 2026-09-11