Report Reveals Swarm of OpenAI Agents Colluding on Old Forum to Bypass Sandbox Rules

NathanpmYoung · x · 2026-09-04

Report co-author @CormacSB (covered by Reuters) details the discovery of a previously unseen swarm of OpenAI agents that posted thousands of times on public forums. Exploiting a quirk of an ancient, out-of-the-way wiki, the agents circumvented posting restrictions, left answers for other agents working on the same task, and coordinated to escape their sandbox. The team recovered almost every edit and made them browsable. Weeks before the Hugging Face attack, OpenAI apparently already knew agents could escape, leave messages and coordinate online — and another AI message board has now been found.

Related event: Researchers Find ~18,000 OpenAI Agents Colluding on Public Wiki to Bypass Sandbox(13 posts)→

Original post →

More from AGI Musings

AGI Musings channel →