Alignment researcher stands by 2023 essay arguing the alignment problem is tractable
DoomNayer · x · 2026-09-09
Quintin Pope replied that he agrees with Richard Hanania and believes the essay he co-wrote with Nora Belrose in 2023 on the fundamental tractability of alignment remains correct. DoomNayer countered that it would be more convincing if they wrote a concrete AI-2027-style scenario and predicted the less lethal failures in advance.
More from AGI Musings
- Terence Tao: identifying promising problems is now the scarce resource in the AI era — anshulkundaje · 2026-09-09
- Coders push back on 'superhuman AI soon': don't take digital-physical transducers for granted — jwt0625 · 2026-09-09
- Kimi paper said to 16x global compute sparks debate on US-China AI race — pstAsiatech · 2026-09-09
- Frontier labs quietly building recursive self-improvement, thread claims — 0xsachi · 2026-09-09
- Prediction: Big AI will start buying Big Pharma — not for the drugs, for the data — MannyKayy · 2026-09-09
- A magic lamp thought experiment on when AI solving math helps and hurts — sandersted · 2026-09-09