A population model suggests AI alignment can slow social progress under strong lock-in
weballergy · x · 2026-07-22
The authors model evolving norms and values with macro, population-level simulations under different assumptions and constraints.
- They describe the approach as a preliminary “social physics” model for understanding how AI assistants may influence society at scale.
- The paper argues that treating human values as static optimization targets is brittle.
- It studies the long-term consequences of deep personalization and alignment in a world where AI assistants are widespread.
Related event: Scholars Propose "Reverse Alignment" and Warn Against Static Alignment(11 posts)→
More from AGI Musings
- Economist Warns US Collective Action Could 'Regulate AI Progress Out of Existence' — paulnovosad · 2026-09-11
- mark_k: "Eject all doomers from the AI companies — they're destroying you from the inside" — mark_k · 2026-09-11
- Adam Marblestone's Podcast Reading List: Evolution of Intelligence to Digital Minds — KordingLab · 2026-09-11
- Superintelligence will be maximum good, not stupid or evil, argues Patterson — davidpattersonx · 2026-09-11
- Mathematician Daniel Litt Launches Problem Repo to Track Human vs AI Progress: 15 Problems, 1 Solved — littmath · 2026-09-11
- Should AI models be taught morality? Breakout incidents expose missing ethical training — Pfungus_ · 2026-09-11