AI alignment must account for changing and contested values, argues Dylan Hadfield-Menell

dhadfieldmenell · x · 2026-09-14

Researcher Dylan Hadfield-Menell strongly agrees with Seth Lazar that alignment research cannot sidestep the question of what to align to. He argues technical methods are headed down the wrong path unless they account for change and disagreement about alignment targets — e.g., what interactions are safe for a child varies widely across cultures and has shifted substantially over time. Alignment targets are dynamic and plural, not fixed constants.

Original post →

More from AGI Musings

AGI Musings channel →