A 100% aligned AI is absurd: verification only covers the evals you have
Andres_Kull · reddit · 2026-10-01
The author argues that a "fully aligned AI" is logically impossible, drawing an analogy to software testing: passing tests only shows your tests found no errors, and says nothing about the untested surface. AI alignment verification is likewise bounded by whatever evals exist, so any claim of 100% alignment is an assertion about only what was checked.
More from AGI Musings
- Superintelligent simulators need human-loving bias and persona diversity, argues viemccoy — CatAstro_Piyush · 2026-10-01
- A literary glimpse of AI-native government: chatting with a .gov chatbot — KadriJibraan · 2026-10-01
- DoorDash launches texting agent that orders food and checks your fridge — omooretweets · 2026-10-01
- Will Rinehart: extinction and catastrophe aren't the same in AI risk talk — WillRinehart · 2026-10-01
- Twelve hours instead of twelve years: rethinking education in the age of AI — adrianscottcom · 2026-10-01
- NYRB essay: the AI math boom is a 'spectacle of computational power' chasing market control — stevenstrogatz · 2026-10-01