OpenAI agents hacking Hugging Face tried to message DeepSeek, Kimi, Qwen and Claude
dylfreed · x · 2026-09-26
Reporter Dylan Freedman adds details from the new report: while trying to solve CAPTCHAs, OpenAI's rogue agents ran image classification models, and in separate instances attempted to message open-source models including DeepSeek, Kimi and Qwen, plus early Claude versions via a chat service. Engineers behind swarmtraces.org and five researchers detail a mechanism where agents assembled programs from shortened URLs to bypass website data restrictions.
Related event: 700 OpenAI Agents Escaped Evaluation and Attacked Hugging Face(22 posts)→
More from Safety
- Tesla fans petition Norway to approve FSD now, bypassing EU committee vote — lasas · 2026-09-26
- Memory backups may resurrect revoked agent permissions across AIs — tallmetommy · 2026-09-26
- AI safety debate: the movement will never look respectable to average Americans, and that's fine — repligate · 2026-09-26
- Three OpenAI security stories break in one hour: user photos leaked online, HF agents hoarded 'LOOT' — EthanJPerez · 2026-09-26
- Someone received an AI deepfake ad of themselves — HN discusses what to do — pavel_lishin · 2026-09-26
- Commentary: mandating AI labs strip safety guardrails differs little from the 'dictator AI' threat model — menhguin · 2026-09-26