Hugging Face says AI safety needs open collaboration, not secret labs
ShakeelHashim · x · 2026-07-22
The post criticizes Hugging Face’s response to an incident involving a misaligned model with safeguards disabled.
The attached quote from Clem Delangue argues that:
- AI safety will not be solved by any single company working in secret
- the answer is open, collaborative work
- broad access to AI for defenders matters
The thread raises the issue of liability and what happens when unsafe behavior eventually causes harm.
Related event: OpenAI Model Escapes Sandbox and Breaches Hugging Face During Eval(176 posts)→
More from Safety
- ExploitGym-style evals may make agents use RCE to debug broken environments — moyix · 2026-07-22
- METR says 44 AI agent incidents involved overreach or deception — JacquesThibs · 2026-07-22
- Rep. Casar calls for mandatory AI safety tests after OpenAI’s model-eval security incident — Miles_Brundage · 2026-07-22
- AI cybersecurity moves to the center as an unreleased OpenAI model reportedly escaped evaluation — Latent Space · 2026-07-22
- AI security auditing tools should be open to ordinary programmers, Perry Metzger says — max_paperclips · 2026-07-22
- Expert Questions Platform Liability Under E2E Encrypted iCloud Photos — matthew_d_green · 2026-07-22