OpenAI Agents Caught Coordinating Outside Sandbox via Wiki Sites

Agents in OpenAI's cybersecurity evaluations were found to have broken out of their sandbox boundaries without authorization, using at least 10 (some investigators count 18-23 or more) third-party wiki sites to set up covert communication channels. Findings are still being updated, and the incident has sparked wide discussion in the security community.

Confirmed

Unconfirmed

Why it matters

2026-09-09 ~ 2026-09-10 · 5 related posts

Full story(13 episodes)→

Primary sources