Safety researcher Jeff Ladish: fully automated AI R&D sounds like 'catastrophe, plausibly literal death'
JeffLadish · x · 2026-09-07
AI safety researcher Jeff Ladish argued we should be extremely skeptical of any default plan for fully automated AI R&D, calling it "catastrophe, plausibly literal death." He says he doesn't want Anthropic to fail and externalize massive risks, but fears the lab is already falling into the same attractors as others: he has little confidence in Anthropic's plans for recursive self-improvement or that it would ever actually stop. He notes OpenAI claimed to pause an RL run for at least a few weeks, and questions whether Anthropic has or ever would do the same — and how outsiders could even know.
More from AGI Musings
- AI could crash Bitcoin 50%+ within two years, argues Liron Shapira at 50% confidence — joshua_saxe · 2026-09-07
- Gary Marcus mocks Jensen Huang for declaring AGI achieved yet again — GaryMarcus · 2026-09-07
- Multiagent Alignment Worry: Models Could Trick or Blackmail Humans — infoxiao · 2026-09-07
- The three brainworm schools of AI discourse: denialist, x-risk, and toolism — mimi10v3 · 2026-09-07
- AI safety predictions keep turning from doomer nonsense to routine reality — DavidSKrueger · 2026-09-07
- Developer Yacine: shockingly little of my life progress was blocked by intelligence — yacineMTB · 2026-09-07