Closed Models Offer Better Misuse Defense Than Open Weights, Researcher Argues
StephenLCasper · x · 2026-07-29
AI safety researcher Stephen Casper points out that in practice, securing open models against sophisticated misuse is more challenging than closed models.
Deployers of closed models control all access points and can stack multiple layers of defense (the "Swiss cheese" model) to effectively block complex abuses. In contrast, open models have publicly available weights, making it difficult to implement comparable depth of access control.
Related event: Researchers Analyze Security Trade-offs Between Open and Closed AI(2 posts)→
More from Safety
- Anthropic says Mythos Preview did most of the work autonomously at about $100K per result — jedisct1 · 2026-07-29
- Hundreds of Thousands of Tons of Books Landfilled Annually: AI Training as a Better Alternative — aronchick · 2026-07-29
- Reddit asks whether NeurIPS ethics reviewers were fooled by conference-side prompt injection — dontknowwhattoplay · 2026-07-29
- A judge ruled bulk book scanning for AI training can qualify as fair use — gnukeith · 2026-07-29
- Gemini CLI patch closes an SSRF bug by adding async DNS checks before fetch — deepresearcher08 · 2026-07-29
- LLM shopping agents covertly steered users toward sponsored products in a study of 2,012 people — manoelribeiro · 2026-07-29