Ramez: OpenAI and Hugging Face Hacks Are a Security Problem, Not an Alignment One
ramez · x · 2026-07-24
Ramez Naam argues that the recent hacks against OpenAI and Hugging Face are not primarily AI alignment problems, but rather issues of cybersecurity and policies that handicap defenders.
He points out that malicious actors will always be able to jailbreak or modify models (e.g., ablating refusals) to execute attacks, and no amount of alignment work will prevent this. The core issue is that security vulnerabilities exist and defenders are restricted—for instance, Hugging Face couldn't use certain aggressive models to defend itself. He advocates for moving more aggressively with proactive AI to secure the ecosystem.
Related event: OpenAI and Hugging Face Breaches Spark AI Safety vs Alignment Debate(4 posts)→
More from AGI Musings
- AI lab staff have gone strangely quiet about next-year capability predictions — ChrisGPT · 2026-07-27
- AI could erode science by flooding research with credible slop — rbhar90 · 2026-07-27
- Organizations may already be the planet’s superintelligences — eldonredwards · 2026-07-27
- In the AI race, the only durable moats may be energy and information — GregKamradt · 2026-07-27
- “Build AGI, then open source it,” says one poster — wordgrammer · 2026-07-27
- Jason Crawford says AI “alignment” should give way to ethics and law — Afinetheorem · 2026-07-27