OpenAI's Former Safety Lead: We Now Ship 'New Capability and Risk Every Tuesday'
sourdub · reddit · 2026-10-08
Former OpenAI safety lead David Robinson told Ezra Klein that reasoning updates, new tools, and coding agents now let OpenAI ship new AI risk weekly.
Key points:
- Previously each new generation started with a massive pretraining run, followed by post-training and safety testing; OpenAI once touted "months" of safety work on GPT-4 after the model was done.
- The base model is now just one layer: pretraining is "the baking," while reasoning training and other post-training steps can be redone quickly—a better reasoning recipe can be layered on without retraining from scratch.
- Release gaps are collapsing: GPT-6.1 Sol arrived at DevDay just a week after GPT-6 Sol.
- "It's not just a chat anymore"—tools and affordances wired into models change capabilities and risks without any new training run.
His warning: safety evaluation cadence hasn't kept up, and risk now compounds on a weekly basis.
More from Companies & People
- MTS live show: OpenAI's 722 math manuscripts, national compute grid, Nvidia deals — ajratner · 2026-10-08
- UT Austin CS opens 2026–27 faculty hiring across tenure-track roles — AkariAsai · 2026-10-08
- Supermemory founder bets his life's work on AI memory as personal brand becomes the last moat — julianweisser · 2026-10-08
- AI podcast Roman Forum hits 200K YouTube subs in 8 episodes, Russell and Bostrom lined up — romanyam · 2026-10-08
- Anthropic, Google, OpenAI and 7 more offer free official AI courses, incl. full MCP track — anthara_ai · 2026-10-08
- Databricks exec: AGI is here, but AI hasn't reached enterprise workflows yet — AnneliesGamble · 2026-10-08