Why AI Alignment Is Hard: The Challenge of Implicit Human Values

_aidan_clark_ · x · 2026-08-06

Discusses the nature and challenges of the AI alignment problem. One side argues that dismissing alignment as a non-issue typically means one hasn't thought deeply about this obviously reasonable yet difficult problem.

The other perspective highlights that humans share value functions to such a high degree that almost everything, even critical requests, is massively underspecified due to an assumed shared resolution of the implicit. Therefore, the core goal of AI alignment is to ensure that AI respects these values as much as the ones humans can explicitly represent.

Related event: Core of AI Alignment Lies in Implicit Shared Human Values(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →