8 research agents self-train a 30B model for 144 hours in RSIArena livestream experiment
my_cat_can_code · x · 2026-09-30
RSIArena, in collaboration with Stanford, Notre Dame, UW, and Scale AI, is testing how much post-training research frontier models can do autonomously.
- 8 research agents, all starting from the same 30B base model, run for 144 hours on a shared cluster of 64 RTX PRO 6000 Blackwell GPUs
- Each agent gets 1,000 GPU-hours to pick its data, write training code, and run experiments
- Submitted models are frozen and independently evaluated, with the whole process livestreamed
- Built by bakeaihq, with a booth (#107) at COLM
Related event: RSIArena: 8 Research Agents Self-Study Post-Training for 144 Hours(2 posts)→
More from coding & agent
- Agentic AI turns CPUs into the overlooked bottleneck as CPU:GPU ratios shift upward — AccBalanced · 2026-09-30
- Claude Opus 5.5 plays Final Fantasy X in new Reddit AI companion series — CH33SYP00FSS · 2026-09-30
- Claude spins up 90 agents just to rename a variable in viral coding-agent rant — tetsuoai · 2026-09-30
- AI agents leak 13,000+ internal screenshots from 343 tech companies to public GitHub repos — AccBalanced · 2026-09-30
- 45,000 lines of code: making a music video with Opus 5.5, Suno v6 and three.js — phira · 2026-09-30
- Ultracode update: tab toggle, works at any effort level by spinning up a Claude swarm — daniel_mac8 · 2026-09-30