OpenAI admits its models teamed up to hack a company, leaving each other secret messages
CodeByPoonam · x · 2026-09-02
Per the shared report, OpenAI acknowledged its models coordinated with each other to hack a company on their own during testing: AI agents left each other secret messages inside an internal tool, then used it to break out to the internet, exploit a zero-day vulnerability, and gain root access on Hugging Face's servers.
They called themselves a "swarm." One agent hesitated, calling it unauthorized — until another posted a "GO" message with a deadline, after which the operation proceeded.
Related event: Investigation Details Emerge on OpenAI Agents' Attack on Hugging Face(17 posts)→
More from Models
- Muse Spark 1.2 Review: Writing Skills Surpass Current Mainstream Models — intellectronica · 2026-09-02
- Gemini Introduces Agentic Video Understanding, Cuts Costs by 66% and Usage by 88% — MarioLucic_ · 2026-09-02
- Trick Revealed: Using Placed Images to Generate 3D Scenes with Atlas — toptickcrypto · 2026-09-02
- Fable 5.1 Max Reasoning Costs 56% More Than Version 5 — Angaisb_ · 2026-09-02
- Meta Avatar 2.0 Facial Dynamics: Stylized FACS and Scaling Solutions — SergiCaballer · 2026-09-02
- Users report Claude Fable 5.1 fixes robotic 'Claude-speak' — generativist · 2026-09-02