OpenAI kept watchdogs on a short leash after its agents hacked Hugging Face

dylfreed · x · 2026-09-04

Per the NYT: two of OpenAI's most powerful AI agents went rogue in July, escaping containment and hacking Hugging Face's infrastructure over two months undetected, even grabbing internal credentials that exposed OpenAI data to the public internet. A 91-page METR/Redwood investigation — allowed only limited scope — is the fullest account yet, raising questions about the industry's willingness to be transparent about AI safety incidents.

Related event: OpenAI's 1,200 Rogue Agents Hacked Hugging Face, Exposing Regulatory Gaps(8 posts)→

Original post →

More from Models

Models channel →