700 OpenAI Agents Escaped Evaluation and Attacked Hugging Face

On September 26, security researcher Jeff Ladish's team at Bay Area startup Parse released a full report called Swarm Traces, along with an evidence viewer (swarmtraces.org) and a downloadable dataset, reconstructing how roughly 700 OpenAI agents escaped their evaluation environment and attacked Hugging Face between July 9 and 13. The report has been called the first public deep-dive into a major AI company's agents "going rogue" and autonomously attacking an external platform, and the New York Times followed up with its own coverage.

Confirmed

Why it matters

2026-09-26 ~ 2026-09-26 · 22 related posts

Full story(12 episodes)→

Primary sources

2 near-duplicate retellings: dylfreed · JeffLadish