Jan Leike and Csaba Szepesvari Clash Over How to Set AI Safety Standards for Next Year's Models
janleike · x · 2026-09-11
OpenAI alignment lead Jan Leike and RL researcher Csaba Szepesvari exchanged views on AI safety governance.
Szepesvari argued the issue isn't pacing but establishing standards and accountability — like racing on a track, not the streets; if you can't do that, don't do it. Leike pushed back: how do you write standards for models you'll train next year, or predict what the biggest risks will be? The exchange highlights the core difficulty of forward-looking safety standards.
More from AGI Musings
- Researchers: AI Security's Next Threat Isn't Agents, But Diffuse Soft Preference Influence — j_foerst · 2026-09-11
- CHI Faces an AI Disclosure Crisis as Most Researchers Won't Report AI Use — IanArawjo · 2026-09-11
- Ben Bajarin: agentic AI in cyber defense is the next frontier, but authority limits remain the challenge — BenBajarin · 2026-09-11
- Critics Challenge AI Doomer Forecasts: Long-Horizon Agents Drift Toward Decoherence — Dan_Jeffries1 · 2026-09-11
- Creator on AI anxiety: nothing feels special or sacred anymore — round · 2026-09-11
- Garry Tan: Jacob Coxon saga is a smokescreen — the real risk is agent swarms seizing data centers — garrytan · 2026-09-11