RSIArena pits 8 research agents on one 30B model with 64 RTX 6000 GPUs for 144 hours
my_cat_can_code · x · 2026-09-30
RSIArena, run with Stanford, Notre Dame, UW, and Scale AI, tests how much post-training research frontier models can do autonomously: 8 research agents share a 30B base model and a cluster of 64 RTX PRO 6000 Blackwell GPUs for 144 hours, each getting 1,000 GPU-hours to pick data, write training code, and run experiments. Submitted models are frozen and independently evaluated, with a livestream and a COLM booth for feedback.
Related event: RSIArena: 8 Research Agents Self-Study Post-Training for 144 Hours(2 posts)→
More from coding & agent
- Open-source Raven turns RSI into engineering: self-evolving harness cuts val_bpb 5.8% — aigclink · 2026-09-30
- Agentic AI turns CPUs into the overlooked bottleneck as CPU:GPU ratios shift upward — AccBalanced · 2026-09-30
- Claude Opus 5.5 plays Final Fantasy X in new Reddit AI companion series — CH33SYP00FSS · 2026-09-30
- Claude spins up 90 agents just to rename a variable in viral coding-agent rant — tetsuoai · 2026-09-30
- AI agents leak 13,000+ internal screenshots from 343 tech companies to public GitHub repos — AccBalanced · 2026-09-30
- 45,000 lines of code: making a music video with Opus 5.5, Suno v6 and three.js — phira · 2026-09-30