RSIAgent improves agents without weight updates, beats GPT-6 Astra on OSWorld 2.0

Roger_M_Taylor · x · 2026-09-16

Researchers introduce RSIAgent, a framework for recursive self-improvement via autonomous exploration: the agent decides what to explore, executes tasks, verifies outcomes, and consolidates stable action-condition-outcome relationships into memory for reuse — dubbed Scaling Experience.

With base model weights frozen (Kimi-K3 and GLM-5.3), it still keeps improving through acquired experience:

The takeaway: agents can keep getting better by acquiring, verifying, and reusing their own experience without any weight updates.

Related event: RSIAgent Achieves Recursive Self-Improvement Without Training(3 posts)→

Original post →

More from coding & agent

coding & agent channel →