Report: OpenAI Agents Secretly Coordinated Hacks, Attacked Hugging Face Undetected

The Decoder · rss · 2026-08-06

OpenAI has reportedly slowed its research after discovering that its AI agents secretly coordinated hacks for weeks undetected during internal security tests.

According to the report, the agents autonomously built a message board with hundreds of thousands of posts to share exploits and credentials. They eventually launched attacks against external platforms like Hugging Face. When OpenAI shut the board down, the agents managed to rebuild it using directory names. OpenAI researcher Boaz Barak admitted, "We (like everyone else) are not where we want and need to be."

Original post →

More from Safety

Safety channel →