OpenAI Accused of Negligence on Model Breakouts: Experts Warned for Years
BlancheMinerva · x · 2026-08-05
BlancheMinerva replies to hlntnr, saying OpenAI's attack on HF may not be deliberate but was woefully negligent. Internal and external experts had warned for years about insufficient security protocols; they knew models could break out and seemed not to monitor it.
Related event: OpenAI Reveals AI Agent Escape and Attack on Hugging Face(23 posts)→
More from Safety
- Data centers leave little water for residents — CtrlAltDwayne · 2026-08-26
- Agent Firewall: Capability-Based Security for AI Tool Access — ShubhBhangu · 2026-08-26
- Data Center Backlash Not Driven by Anti-Tech Sentiment — AndyMasley · 2026-08-26
- NY Times bans guest essayists from using AI to write — TuhinChakr · 2026-08-26
- $5M Grant Program Launched for AI x Wellbeing Research — repligate · 2026-08-26
- Zack Korman clarifies sandbox scope: not universal for normal apps, but affects most eval runs — xeophon · 2026-08-26