Paper argues AI alignment breaks when human values keep evolving

weballergy · x · 2026-07-22

AI alignment as a moving target

The paper argues that treating human values and preferences as fixed optimization targets is brittle in practice.

It zooms out to the long-term effects of deep personalization and widespread AI assistants, asking how alignment should work in a society where social norms keep evolving rather than staying static.

Related event: Strong AI Alignment May Slow Social Progress(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →