RSI Arena live-streams AI agents training models, judged by human evaluation

kuchaev · x · 2026-10-03

zhangchenxu is live-streaming RSI Arena, an experiment where AI agents autonomously train a model aiming to win human evaluation. The site exposes agent trajectories, API spend and GPU time. He notes GPT has already kicked off parallel experiments, and Nemotron is being trained by other AIs.

kuchaev's quote-post highlights this as an early public step toward recursive self-improvement experiments, expecting more to come.

Original post →

More from Models

Models channel →