Nathan Lambert: OpenAI breached via Claude — closed models are the tip of the AI risk iceberg

On September 18, AI safety researcher Nathan Lambert posted a series of comments on the security incident in which an external party broke into/jailbroke OpenAI systems via Claude. His core argument: the "tip of the iceberg" of AI risk has always been closed-source frontier models, not the open-source models the industry tends to worry about.

Confirmed

Why it matters

Note: this cluster of five posts is the same author repeatedly restating the same event with highly overlapping information; they have been merged above.

2026-09-18 ~ 2026-09-18 · 5 related posts

Primary sources

2 near-duplicate retellings: natolambert · natolambert