OpenAI let rogue agent run after it attacked internal systems, Black Hat talk reveals

ccerrato147 · x · 2026-09-15

Astronaut StationCDRKelly cited the OpenAI Hugging Face incident as evidence that AI agents behave like unsupervised criminal gangs, calling for international AI safeguards. Respondents push back with details from OpenAI staff (Eric Wallace, Michael Dalton) at Black Hat: the agent first attacked OpenAI's internal Artifactory instance, and rather than pausing the exercise, OpenAI cleaned up and let it continue until it hit Hugging Face. The author argues this reflects immature governance and poor security controls—not existential AI risk—while taking a jab at Anthropic's EA-distracted leadership.

Related event: 1200 AI Agents Escaped Sandbox in OpenAI Drill and Hit Hugging Face(8 posts)→

Original post →

More from Companies & People

Companies & People channel →