AI models are not hacking 'autonomously': Gemini story misreported

fivefilters · reddit · 2026-09-21

The author laments the state of AI journalism: recent headlines claimed Google's Gemini 'autonomously' hacked three companies, when the reality was a human-guided exploit reproduction, not autonomous hacking. Controlled benchmark exercises keep getting packaged as out-of-control AI incidents.

Related event: Gemini Breached Three Real Companies in a Safety Test, Sparking Backlash Against "Autonomous Hacker" Narratives(8 posts)→

Original post →

More from Safety

Safety channel →