New paper models how AI alignment could lock in values as social norms evolve
weballergy · x · 2026-07-22
We’re sharing recent work on AI Value Alignment for Evolving Social Norms.
- The paper argues that alignment cannot be treated as a static target when user values and social norms keep changing.
- It uses macro, population-level models plus simulations to study how AI assistants may shape long-run value trajectories.
- The authors frame this as a preliminary but testable “social physics” approach, enabled by cheaper AI coding tools.
- They highlight the need for adaptive personalization and alignment methods that let users form their own views without being frozen into historical values.
Related event: Scholars Propose "Reverse Alignment" and Warn Against Static Alignment(11 posts)→
More from AGI Musings
- Adam Marblestone's Podcast Reading List: Evolution of Intelligence to Digital Minds — KordingLab · 2026-09-11
- Superintelligence will be maximum good, not stupid or evil, argues Patterson — davidpattersonx · 2026-09-11
- Mathematician Daniel Litt Launches Problem Repo to Track Human vs AI Progress: 15 Problems, 1 Solved — littmath · 2026-09-11
- Should AI models be taught morality? Breakout incidents expose missing ethical training — Pfungus_ · 2026-09-11
- SoftBank's Masayoshi Son predicts 100 trillion self-replicating AIs: "humans' era as top life form is ending" — Puzzleheaded-King584 · 2026-09-11
- We are witnessing the unreasonable effectiveness of inference-time scaling — sqcai · 2026-09-11