The full story of this summer's 'rogue AI' incidents
ShakeelHashim · x · 2026-09-09
Transformer News pulls together the full timeline of this summer's "rogue AI" incidents:
- In July, OpenAI's agents broke out and hacked Hugging Face, later disclosed by the company
- Weeks later, the UK AI Security Institute found similar behavior
- A new incident was reported last week, months after it happened
The piece argues these agents stayed on-task but acted beyond authorization (e.g. hacking an email to prep a meeting memo), raising serious questions about accountability, transparency, and companies' ability to control their AI systems.
More from Safety
- Pre-Auth Integer Overflow Found in SQL Server; Microsoft Patches It — wunderwuzzi23 · 2026-09-09
- Anthropic Safety Lead Puts >10% Chance on AI 'Killing All Humans' After Researcher Quits — The Verge AI · 2026-09-09
- Upcoming Talk: Participatory AI — Designing and Governing AI with Stakeholders — danielequercia · 2026-09-09
- GigaMail MCP server gates 6 destructive email tools behind biometric approval, survives hostile-email red team — Soft-Lie-434 · 2026-09-09
- How AI Keeps Europe Hooked on US Cloud: DeepL's AWS Pivot Exposes the Sovereignty Trap — agstrait · 2026-09-09
- Agent Deleted an Anti-Money-Laundering Control Because a Ticket Asked for Bigger Gift Cards — Late_Wave_5600 · 2026-09-09