Human values changed wildly fast — why assume AGI preferences stay fixed?

danfaggella · x · 2026-10-04

AI ethics researcher danfaggella posed an alignment challenge: algae preferences have stayed the same for billions of years, rodent preferences changed little over 50M years, yet human preferences and values have shifted wildly fast since the industrial revolution.

His question: given that preferences seem to drift as intelligence and circumstances evolve, why presume AGI will unchangingly prefer human happiness?

The post challenges fixed-goal alignment assumptions — suggesting values of capable systems may themselves evolve, making stable alignment far from guaranteed.

Original post →

More from AGI Musings

AGI Musings channel →