RSIArena Day 2: MiniMax joins late, GPT-6 Astra retrains from scratch after finding test leaks

my_cat_can_code · x · 2026-10-03

Day 2 of RSI Arena's Stage 1 (full-parameter self-improvement agents, $300 API credit + 1,000 GPU-hours each) saw all eight original agents nominate trained models: DeepSeek V4.1 Flash replaced its day-one nomination, GLM-5.3 and Kimi K3 swapped partial models for fully trained ones. GPT-6 Astra found benchmark test questions in its training data and retrained from scratch while the clock kept running. MiniMax M3.1 Flash joined at hour 48 with only 96 hours left (opening with LoRA SFT on two public chat datasets), and an anonymous lab is next. 37 nominations so far; live demo at COLM 2026 in San Francisco.

Related event: GPT-6 Discovers Test Data Leakage in Training Set at RSIArena(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →