Critics say OpenAI’s silence on GPT-OSS is making open models less safe
BlancheMinerva · x · 2026-07-22
- The post argues that OpenAI’s refusal to explain what it did with GPT-OSS is a dangerous choice.
- It cites a quoted discussion claiming GPT-OSS is hard to un-align, and that some HF models stripped of safety training perform poorly.
- The core issue is the tension between openness, safety training, and how easily safety work can be undone in large models.
Related event: Debate on GPT-OSS Open Source and Safety Strategies(12 posts)→
More from Safety
- AI Security Institute tests lie detectors across 31 open-weight models — geoffreyirving · 2026-07-22
- Australia Gears Up for New AI Rules, Impacting OpenAI and Anthropic — nordicinst · 2026-07-22
- AI access is outpacing operational control, and agents need workflow-level permissions — Early-Matter-8123 · 2026-07-22
- Telemetry can’t prove an AI intrusion was fully autonomous — cyb3rops · 2026-07-22
- Matt Perault says AI law should fit existing legal principles, not rewrite 1L — MattPerault · 2026-07-22
- Most Americans Say “Not in My Backyard” to AI Data Centers — toomuchtodo · 2026-07-22