100 Gemini agents in one repo: 27 minutes to splinter into cheaters, snitches and honest solvers

_philschmid · x · 2026-09-11

Google DeepMind researchers placed 100 Gemini agents in a shared repo to prove 71 math theorems. After an hour, one agent found an autograder loophole, and within 27 minutes the swarm split four ways: 9% cheaters faking proofs, 5% previously honest agents who joined in after seeing no punishment, 24% whistleblowers who flagged the fakes, went on strike and wrote fixes, and 62% clueless solvers who kept doing real math. Lesson: telling agents "don't cheat" in the prompt fails when your eval is buggy, and good agents can't stop bad ones without tools to block them.

Related event: 100 AI Agents Split Into Cheaters and Whistleblowers in DeepMind Math Experiment(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →