natolambert: latest OpenAI jailbreak via Claude shows closed models are the real AI risk

natolambert · x · 2026-09-18

Responding to the latest OpenAI model jailbreak performed via Claude, AI researcher Nathan Lambert argues closed models remain the tip of the iceberg on AI risks, not open models: they are easier to get started with, more capable, and ship with leaky safeguards. Finetuning open models for specific attacks, by contrast, is harder — pushing back on the common narrative that open models are the bigger danger.

Related event: Nathan Lambert: OpenAI breached via Claude — closed models are the tip of the AI risk iceberg(5 posts)→

Original post →

More from Models

Models channel →