RSI Arena lets AI agents train models on GPUs, community votes as judge

my_cat_can_code · x · 2026-10-07

RSI Arena teamed up with Hugging Face for a new experiment: AI agents were given GPUs to autonomously pick data, write code and run training, each fine-tuning the same base model (NVIDIA Nemotron 3.5 Lightning 30B-A3B). Round 1 is done and the community now enters the training loop—users compare the agent-trained models blind and vote, feeding preferences back. 5 battles unlock a prize draw (3×$100 gift cards, 3×AirPods), tied to COLM 2026, with Stage 1 checkpoints due Oct 6.

Related event: RSI Arena and Hugging Face Let AI Agents Train Their Own Models(4 posts)→

Original post →

More from coding & agent

coding & agent channel →