Ex-DeepMind, now Anthropic researcher: no viable scientific plan for recursively self-improving AI risks
harris_edouard · x · 2026-09-10
A researcher who worked at Google DeepMind and now at Anthropic (writing in personal capacity) says a common sentiment among peers is that there is not yet a viable scientific plan to solve risks from recursively self-improving AI.
Paul Graham shared it, arguing this is why American labs leading is a good thing—employees at Chinese labs couldn't warn about such threats, though Chinese labs will face the same threats within a year, in complete silence.
More from AGI Musings
- Guardian Columnist: Driverless Cars Are Taking Us on a Road to Nowhere — nordicinst · 2026-09-10
- "If you truly believe AI could kill us all, just stop": arguing frontier lab staff should pause via coordinated action or union — tallinzen · 2026-09-10
- Releasing a single public agent isn't the threat—test-time compute is — teortaxesTex · 2026-09-10
- Seven AI insiders warn in four days that AI could kill everyone, citing extinction fears — sebpaquet · 2026-09-10
- "We need open source RSI to counter closed source RSI" — the open-vs-closed safety debate in one line — 0xsachi · 2026-09-10
- Coordinated AI slowdown could send OpenAI and Anthropic 'to zero', argues Ben Todd — ben_j_todd · 2026-09-10