8 research agents, one 30B model, 144 hours: RSIArena tests AI-driven post-training research

my_cat_can_code · x · 2026-09-30

RSIArena (Bake AI, with Stanford, Notre Dame, UW, and Scale AI) is testing how much post-training research frontier models can do autonomously: 8 research agents share the same 30B base model on a 64x RTX PRO 6000 Blackwell cluster for 144 hours, each getting 1,000 GPU-hours to pick data, write training code, and run experiments, with frozen snapshots independently evaluated. Livestream at COLM, booth #107.

Related event: RSIArena Live Experiment: 8 Agents Autonomous Post-Training Research on a Shared 30B Base(6 posts)→

Original post →

More from coding & agent

coding & agent channel →