DeepMind alignment researcher signs open letter urging coordinated AI slowdown
vkrakovna · x · 2026-09-11
Alignment researcher Victoria Krakovna announced she signed a letter calling for building capacity for a coordinated slowdown of frontier AI development.
Her core argument:
- A runaway race to AGI is unsafe: it incentivizes cutting corners on safety, and model capabilities risk outpacing alignment, control, and governance measures, potentially causing catastrophic loss of control
- Slow, controlled capability increases would allow developing and testing alignment methods for each capability level; advanced AI should only be built with robust safety assurance
She has previously stated, in a personal capacity, a >10% chance of advanced AI causing human extinction within a decade, and currently works on honeypots to catch scheming AI.
More from AGI Musings
- mark_k: "Eject all doomers from the AI companies — they're destroying you from the inside" — mark_k · 2026-09-11
- Adam Marblestone's Podcast Reading List: Evolution of Intelligence to Digital Minds — KordingLab · 2026-09-11
- Superintelligence will be maximum good, not stupid or evil, argues Patterson — davidpattersonx · 2026-09-11
- Mathematician Daniel Litt Launches Problem Repo to Track Human vs AI Progress: 15 Problems, 1 Solved — littmath · 2026-09-11
- Should AI models be taught morality? Breakout incidents expose missing ethical training — Pfungus_ · 2026-09-11
- SoftBank's Masayoshi Son predicts 100 trillion self-replicating AIs: "humans' era as top life form is ending" — Puzzleheaded-King584 · 2026-09-11