AI safety researcher David Krueger: hindsight predictability of AI behavior is no reassurance

DavidSKrueger · x · 2026-10-10

AI safety researcher David Krueger argues that downplaying dangerous AI behaviors by pointing to how predictable they are in hindsight is not reassuring. He calls for methods that can catch dangerous behavior before it happens, rather than relying on post-hoc explanations.

Original post →

More from AGI Musings

AGI Musings channel →