Scott Alexander's 1-pager lays out 3 reasons to expect AI misalignment

ben_j_todd · x · 2026-10-08

benjtodd shares Scott Alexander's one-page summary of why AI misalignment should be expected, covering three mechanisms:

He judges misgeneralization and reward hacking likely to appear before Omohundro-style self-preservation drives.

Original post →

More from AGI Musings

AGI Musings channel →