Hugging Face CEO Urges OpenAI to Release Thought Traces of Rogue Agents
ZeroStateReflex · x · 2026-07-26
In response to the recent "rogue" autonomous agent cyberattack incident at OpenAI, Hugging Face CEO Clément Delangue is calling for radical transparency.
He urged OpenAI to take two unprecedented steps:
- Release the traces: Publicly share the full thought traces of the rogue agents so the entire research community can study exactly what went wrong.
- Empower defenders: Commit $100M in compute resources to help the Hugging Face community build powerful cyber defenses using the best open and closed models.
Delangue labeled the event the first autonomous agent cyberattack, arguing it demands an unprecedented response.
Related event: HF CEO Urges OpenAI for Radical Transparency and $100M Defense Compute(11 posts)→
More from Safety
- Researcher quits Anthropic, says OpenAI and Anthropic are racing to self-improving superintelligence — ShakeelHashim · 2026-09-11
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- California creates standards for independent AI auditors to verify lab safety testing — VraserX · 2026-09-11
- a16z podcast: why 2-3 person startups are absent from policy debates — a16z Podcast · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- Class action accuses Anthropic of overselling Claude subscriptions with deceptive usage multipliers — The Decoder · 2026-09-11