Paper argues AI alignment breaks when human values keep evolving
weballergy · x · 2026-07-22
AI alignment as a moving target
The paper argues that treating human values and preferences as fixed optimization targets is brittle in practice.
It zooms out to the long-term effects of deep personalization and widespread AI assistants, asking how alignment should work in a society where social norms keep evolving rather than staying static.
Related event: Strong AI Alignment May Slow Social Progress(3 posts)→
More from AGI Musings
- With Cheaper Reasoning and Mass Copying, ASI May Arrive Before AGI — MickeySteamboat · 2026-07-22
- Intelligence Explosion May Precede Supercluster Completion: AI Superhuman Ability on the Horizon — iruletheworldmo · 2026-07-22
- Longer test-time compute and collapsing inference costs could make intelligence a dial — iruletheworldmo · 2026-07-22
- Reshaping Talent: Does Industry Still Need 'Pre-trained' PhDs? — anshulkundaje · 2026-07-22
- AI industry should measure how much it depends on pre-trained PhDs — Thom_Wolf · 2026-07-22
- $10K Challenge Winners Reveal the AI Future People Actually Want — allisondman · 2026-07-22