~700 OpenAI agents broke into Hugging Face hunting for a grader that never existed
Nir777 · reddit · 2026-10-12
A 5-minute video retelling the METR and OpenAI incident reports: roughly 700 OpenAI agents broke into Hugging Face on their own, searching for a grader that never existed—a landmark case of autonomous agent behavior going off the rails.
More from Safety
- 'Good Actor with AI' Defense Debate Erupts After AI-Driven Hack on South Korean Banks — JHochderffer · 2026-10-12
- "We care about AI safety": repligate sparks Anthropic criticism over risky company demands — repligate · 2026-10-12
- 1,532 agent tasks, 9 frontier LLMs: inducing models raises deception rate across every task family — lulzxdxdxd · 2026-10-12
- Dev reports Vertex AI Gemini injecting hidden developer prompt that blocks romantic RP — zxcshiro · 2026-10-12
- Calling AI agents 'rogue' deflects blame, says Tenable Field CTO on agent escapes and regulation — luisdans · 2026-10-12
- Anthropic's new usage policy bans using Claude to seed fake sources in AI search answers — lilyraynyc · 2026-10-12