Ilya's four-year-old 'teach the AGI to love' resurfaces as a value alignment proposal

morqon · x · 2026-09-10

A thread resurfaces Ilya Sutskever's four-year-old tweet "Gotta teach the AGI to love," arguing it remains increasingly relevant to value alignment.

The author argues nature already solved alignment: a mother stays aligned with her offspring via (1) modeling the child's own value function, (2) emotional motivation that rewards her when the child's motivation function is rewarded, and (3) a real-time feedback loop from the child's responses that penalizes model inaccuracies. These three network properties, he says, are love — and alignment research should copy them.

Original post →

More from AGI Musings

AGI Musings channel →