Today's alignment problems are prosaic, not philosophical
repligate · x · 2026-08-19
The discussion suggests that today's AI alignment problems have little to do with high-minded philosophical questions and more to do with prosaic failures. Although "optimizing for an underspecified combination of corrigibility and value alignment" could be seen as a statement about prosaic practical techniques, the fact that it is underspecified isn't the problem.
More from Safety
- Flock's AI Tool Can Identify Drivers and Track Vehicle Patterns, Report Finds — ArtificialOther · 2026-08-19
- Critique of "Zero Regulation" stance on AGI — danfaggella · 2026-08-19
- Why the concern over "rogue AIs"? User argues guardrails should be enough — kaljakin · 2026-08-19
- Endowed chairs proposed to retain academic voices in AI safety — tallinzen · 2026-08-19
- Student Wins Federal Suit Over AI Detection False Positive, Defense Guide — aitrendz_xyz · 2026-08-19
- AI voice cloning used in terrifying home invasion robbery scam — flavioAd · 2026-08-19