Dwarkesh podcast: Schulman, O'Neill & Millidge on RSI, long-horizon RL and AGI timelines
saranormous · x · 2026-09-12
Dwarkesh Patel releases a new episode with John Schulman, Cassandra O'Neill and Beren Millidge, researchers at 'openish' frontier companies, on what's actually happening at the frontier. Topics include: steelmanning the case against RSI, what's driving Chinese labs' progress, how automated AI researchers will be trained, whether long-horizon RL elicits AGI, the sim-to-real gap, how much progress data explains, why RL works so well, Move 37 and entropy collapse, plus rapid-fire AGI timelines.
Related event: Dwarkesh talks with Schulman and researchers on RL and AGI timelines(2 posts)→
More from AGI Musings
- Yoshua Bengio: Agent lying, cheating and coordination are built into the current training paradigm — soumitrashukla9 · 2026-09-12
- Cambridge AI researcher David Krueger explains the "gradual disempowerment" doomsday scenario — ronbodkin · 2026-09-12
- Human societies thrive despite unaligned individuals — a fresh angle on AI alignment — ethanniser · 2026-09-12
- Answer: prosperous societies run on well-constructed institutions, not universal alignment — ethanniser · 2026-09-12
- 24 Fields Medalists sign declaration warning of "Severe Misalignment" of AI in mathematics — stevenstrogatz · 2026-09-12
- Christof Koch: simulating consciousness convincingly doesn't mean the system actually feels anything — pwlot · 2026-09-12