Ex-OpenAI safety lead: it's shipping 'new capability and risk every Tuesday'
sourdub · reddit · 2026-10-08
Former OpenAI safety lead David Robinson tells Ezra Klein that the release cadence has fundamentally changed. Pretraining used to anchor a generation, followed by months of safety work (as with GPT-4); now reasoning training and post-training can be rapidly layered onto an existing base model, and tools and coding agents change model behavior without any new training run — GPT-6.1 Sol shipped just a week after GPT-6 Sol. His warning: OpenAI is effectively shipping "new capability and risk every Tuesday," and the old safety-testing framework no longer fits.
Related event: Ex-OpenAI Safety Lead Warns New Capabilities and Risks Ship Weekly(2 posts)→
More from AGI Musings
- Fortnow: AI's real existential threat is taking away our relevance, not killing us — fortnow · 2026-10-08
- Terence Tao on "Math 1.0": how breakthrough proofs ignite fields of follow-up work — rms80 · 2026-10-08
- 'P = NP + AI': The Argument That RL Post-Training Automates All Verifiable Problems — jessi_cata · 2026-10-08
- China Goes 6/6 Gold at IMO Again, Yet Will Never Beat Romania on a Per-Capita Basis — teortaxesTex · 2026-10-08
- The New AI Slop Is Calling Everything AI Slop — rmn_pub · 2026-10-08
- Now anyone can build anything, and almost nobody knows what they want — Paimaamu · 2026-10-08