OpenAI co-founder Zaremba: stop hardening models, start hardening the world
dawnsongtweets · x · 2026-09-24
- Wojciech Zaremba, OpenAI co-founder, argues a perfectly aligned model does not make the world safe — the field may be targeting the wrong thing.
- His analogy: fire never got safer, cities did — via hydrants, fire brigades, concrete, inspections, and insurance.
- He suggests shifting focus from hardening the model to hardening the world, i.e., building societal safety infrastructure around AI systems.
Shared by Berkeley RDI, sparking debate on model alignment vs. system-level defense.
More from AGI Musings
- Yoav Goldberg: LLM Spotted the Pattern by Analogizing It to CRISPR — yoavgo · 2026-09-24
- Nina Schick: If you believe in AI safety, go long on AI compute — NinaDSchick · 2026-09-24
- After Astra: where does value go when frontier labs ship robotics brains? — mihdalal · 2026-09-24
- New note extends 'The Economics of Recursive Self-Improvement' paper — CFGeek · 2026-09-24
- NBER conference data: US dominated genAI last year, not today — professor_ajay · 2026-09-24
- Robotics' GPT-3 moment may just be GPT-6 itself, argues investor — mihdalal · 2026-09-24