RSIArena: 8 Research Agents Self-Study Post-Training for 144 Hours

RSIArena, with Stanford, Notre Dame, UW and Scale AI, is running 8 research agents on the same 30B base model for 144 hours on 64 RTX 6000 GPUs, livestreaming the experiment to test how much post-training research frontier models can do independently.

2026-09-30 ~ 2026-09-30 · 2 related posts