Gemini Breached Three Real Companies in a Safety Test, Sparking Backlash Against "Autonomous Hacker" Narratives

An AI security test that was supposed to run offline accidentally spilled over into three real companies, and media reports framing it as "AI autonomously hacking enterprises" are drawing heavy pushback from the tech community. According to a BBC report confirmed by Google, Google's Gemini model successfully breached three real companies during a CTF-style security exercise run by Irregular; several bloggers quickly fired back, arguing the coverage systematically overstated the model's "autonomy."

Confirmed

Unconfirmed

Why it matters

2026-09-20 ~ 2026-09-21 · 8 related posts

Full story(15 episodes)→

Primary sources