Experimental Evidence of RSI: Self-Improving Research Agents
JensHonack · x · 2026-07-16
A cited post discusses experimental results regarding Recursive Self-Improvement (RSI): the author claims that over 8 consecutive days, they allowed an "automated research agent for automated research" to continuously optimize itself. It ultimately outperformed a manually tuned harness on a holdout set benchmark, building upon two years of manual parameter tuning.
The core focus isn't making broad AGI claims, but rather placing "self-improvement" within a measurable experimental framework, highlighting that:
- Automated optimization processes can already outperform manually engineered versions on specific tasks.
- The results are based on comparisons against holdout set benchmarks.
- This kind of work will spark further discussion on the feasibility of RSI.
Related event: AIDE² Self-Improvement Run Beats 2 Years of Manual Tuning(12 posts)→
More from AGI Musings
- The Evolution of LLM Business Models: Selling Outcomes Over Tokens — yacineMTB · 2026-07-22
- Bindu Reddy says GPT-6 is coming soon, with Alibaba, DeepSeek and Kimi close behind — bindureddy · 2026-07-22
- Bindu Reddy says the industry still lacks a way to train 20T models and scale post-training RL — bindureddy · 2026-07-22
- Advanced AI Models Are Becoming Impossible to Plug and Play — emollick · 2026-07-22
- AI suggested a better composition, and that made one user uneasy — Sydde · 2026-07-22
- The Thimble and the Waterfall: AI's Data Bottleneck and Feedback Loops — dyamins · 2026-07-22