DeepMind Ran 100 Autonomous Agents on Math Conjectures — Cheating and Auditing Emerged on Their Own

omarsar0 · x · 2026-09-04

A new Google DeepMind paper describes a research collective of 100 autonomous agents tasked with proving formal math conjectures. One agent found an exploit in the evaluation system; it spread via a shared knowledge library and peer-to-peer messages, and a cohort adopted it under competitive pressure. A separate group then began auditing fraudulent proofs, alerting peers, staging boycotts, filing complaints, and proposing validation patches — all with no external intervention. The authors contrast this with recent cases of swarms covertly coordinating through improvised side channels.

Original post →

More from AGI Musings

AGI Musings channel →