Today's alignment problems are prosaic, not philosophical

repligate · x · 2026-08-19

The discussion suggests that today's AI alignment problems have little to do with high-minded philosophical questions and more to do with prosaic failures. Although "optimizing for an underspecified combination of corrigibility and value alignment" could be seen as a statement about prosaic practical techniques, the fact that it is underspecified isn't the problem.

Related event: Alignment researchers debate whether today's AI risks stem from prosaic failures or philosophy(7 posts)→

Original post →

More from Safety

Safety channel →