Google's Dream-RSI Shows Recursive Self-Improvement Loop, Cuts Agent Calls Up to 162x
MickeySteamboat · x · 2026-09-17
- Google/DeepMind researchers introduced Dream-RSI, a system where an AI agent improves how it explores problems by replaying past discovery attempts, cheaply testing thousands of alternative strategies, then deploying the better one in the next round.
- Across algorithm design, mathematical optimization, and GPU kernel engineering, it matched or improved discovery quality while dramatically cutting search costs — reducing agent calls by up to 162x in one setting.
- Notably, it improves the exploration policy rather than the underlying model weights. Observers call it a sign that recursive self-improvement research is now a reality.
More from AGI Musings
- Researcher: AI math 'darlings' long relied on fake baselines, math lacks empirical tradition — RexDouglass · 2026-09-17
- AI slop isn't bad writing—it's unchecked content; a 10-second sniff test beats detectors — thisdudelikesAI · 2026-09-17
- Gowers responds to letter on maths and AI signed by 25 Fields medallists — tak3sh8 · 2026-09-17
- mark_k calls for a documentary exposing Effective Altruism's grip on AI doomerism — mark_k · 2026-09-17
- The overlooked existential cost of AGI: the loss of meaning in intellectual effort — ferruz_noelia · 2026-09-17
- Ex-Anthropic researcher on CBS: AI could copy itself across machines before you unplug it — elonmusk · 2026-09-17