David Krueger: four unresolved foundational problems stand between us and safe AI

DavidSKrueger · x · 2026-09-22

Cambridge's David Krueger lays out four core gaps in AI safety research: we don't know how AI systems work (interpretability), can't predict their behavior (testing), can't stop misbehavior (alignment), and can't ensure control if it happens (control). He argues foundational questions in all four fields remain unresolved despite years of effort, so we cannot rely on solving them to a deadline.

Original post →

More from Safety

Safety channel →