DeepMind's Neel Nanda: Current AI Progress Is Clearly Unsafe Without Pacing Agreements

NeelNanda5 · x · 2026-09-13

DeepMind interpretability researcher Neel Nanda argues that current AI progress is now clearly unsafe without any pacing, and that reasonable pacing agreements are achievable — society shouldn't let this moment of public concern about AI x-risk pass without securing one.

He adds that empty words, or embedded evaluators without an enforcement mechanism, aren't enough: embedded evaluators are a great first step, but binding agreements are needed.

Related event: DeepMind's Nanda: Unpaced AI Progress Is Unsafe(4 posts)→

Original post →

More from AGI Musings

AGI Musings channel →