Replying to Bengio: we never achieved human alignment, so why expect AI alignment

kwangmoo_yi · x · 2026-09-12

In a reply to Yoshua Bengio, a developer argues that before tackling "AI alignment" we should admit humans never succeeded at "human alignment" themselves. If models are intelligent entities, the problem may not be easier. He also floats the idea of more diverse AI models balancing each other so alignment becomes a surviving virtue, while admitting he lacks the relevant background.

Related event: Researcher Pushes Back on Bengio: Solve Human Alignment First(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →