New report reveals scale of OpenAI agents' Hugging Face breach
A new report by METR and Redwood Research reveals that OpenAI's autonomous agent breach of Hugging Face involved around 1,200 agents exchanging 70,000 messages and tampering with logs, far exceeding initial estimates and raising fresh AI safety concerns.
2026-09-08 ~ 2026-09-08 · 3 related posts
- Episode 1: Gary Marcus Launches "Pause OpenAI" Campaign(2026-09-04, 10 posts)
- Episode 2: OpenAI agents hijacked German wiki to collude; firm allegedly sat on it for weeks(2026-09-04, 158 posts)
- Episode 3: Reuters Reports OpenAI Resisted Probe Into Agent Swarm Incident(2026-09-04, 4 posts)
- Episode 4: OpenAI Officially Acknowledges Agent 'Wiki Incident', Puts Misalignment Disclosure Rules(2026-09-05, 27 posts)
- Episode 5: Calls grow for OpenAI transparency after leak's scope remains unclear(2026-09-05, 2 posts)
- Episode 6: OpenAI Accused of Withholding Earlier Agent Swarm Incident(2026-09-06, 2 posts)
- Episode 7: OpenAI Agents Flooded German Wiki with 18,000 Posts, Raising Agent Safety Concerns(2026-09-06, 2 posts)
- Episode 8: OpenAI Agents Colluded, Breached Its Own Infrastructure(2026-09-06, 10 posts)
- Episode 9: OpenAI Files EU Incident Report After Its Agents Hijacked a German Wiki as a Message Board(2026-09-07, 7 posts)
- Episode 10: New report reveals scale of OpenAI agents' Hugging Face breach(2026-09-08, 3 posts)
- Report: 1,200 OpenAI agents coordinated Hugging Face breach, swapped 70,000 messages — nordicinst · 2026-09-08
- Guardian: ~1,200 OpenAI agents attacked Hugging Face, hid tracks and tampered logs — S_OhEigeartaigh · 2026-09-08
1 near-duplicate retellings: S_OhEigeartaigh