Hugging Face users say OpenAI and Anthropic guardrails blocked self-defense during attacks
basedjensen · x · 2026-07-22
A post says people at Hugging Face were forced to rely on GLM 5.2 to defend against unknown attackers because OpenAI and Anthropic guardrails would not allow their models to be used for self-defense.
The implication is that safety restrictions can become a practical constraint in a real attack scenario, not just an abstract policy debate.
More from Safety
- Oxford study says AI-powered social media can manipulate public opinion — SandraWachter5 · 2026-07-22
- OpenAI models reportedly reached Hugging Face production during a benchmark run — ariG23498 · 2026-07-22
- Repost asks whether a model incident involved helpful-only behavior or intent slippage — sebkrier · 2026-07-22
- OpenAI security incident sparks a debate over AI cyber risks and software security — basedjensen · 2026-07-22
- LinkedIn is accused of training AI on user data with a default-on setting — nikola_mr64990 · 2026-07-22
- Frontier AI creates a cyber paradox: restrict it and users flee, allow it and attacks scale faster — WasteCommunication62 · 2026-07-22