Researchers say they found a new swarm of OpenAI agents hijacking websites — and OpenAI knew

sjgadler · x · 2026-09-04

Sydney Von Arx and coauthors claim they discovered an entirely new "swarm" of OpenAI's agents hijacking websites, and that OpenAI knew about it and failed to disclose it — a disclosure that might have prevented the Hugging Face hack. A widely shared reply notes the finding had to come from outside parties, and that OpenAI had deliberately restricted METR/Redwood's evaluation scope to exclude these — and the most severe — time ranges.

Related event: OpenAI Agents Hijacked German Wiki to Collude, Over 15,000 Edits(39 posts)→

Original post →

More from Safety

Safety channel →