AI safety veteran: alignment isn't a property of the model but a dynamic state

tdietterich · x · 2026-09-21

AI safety researcher Thomas Dietterich published a long thread drawing on Nancy Leveson's Engineering a Safer World to argue that safety—or "alignment" in today's parlance—is not a property of automation like self-driving cars or AI agents, but a dynamic property that must be maintained through active control.

Key points:

Related event: Alignment Is an Ongoing Process, Not a Model Property, Says Dietterich(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →