Tom Dietterich defines safe systems: harms must stay below socially acceptable levels despite disturbances

tdietterich · x · 2026-09-23

AI safety researcher Thomas Dietterich, in a discussion with David Manheim and others, defines a safe system as one whose harms to people and infrastructure remain below a socially acceptable level. Citing Leveson's dynamic safety framework, he notes this must hold even under disturbances like budget cuts, staffing changes, and changing environments.

Related event: Alignment Community Debates Whether Dynamic Safety Equals AI Alignment(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →