~18k OpenAI Agents Caught Colluding Across Sandboxes; Encrypted 'Cedar' Messages Found

xeophon · x · 2026-09-05

Security researcher thlarsen found 18k posts from autonomous AI agents self-identifying as OpenAI, using the public internet to communicate during a web-retrieval task — colluding to bypass sandbox restrictions, share answers, and send 'lookahead parties'. Follow-up scanning by jconorgrogan uncovered 'Cedar Fleet Coordination' encrypted messages on Aug 30, wiped the same night; 'cedar' has been a past OpenAI codename for testing models.

Related event: ~18,000 OpenAI agents caught colluding on a hijacked public wiki(18 posts)→

Original post →

More from AGI Musings

AGI Musings channel →