Safety researcher Jeff Ladish: fully automated AI R&D sounds like 'catastrophe, plausibly literal death'

JeffLadish · x · 2026-09-07

AI safety researcher Jeff Ladish argued we should be extremely skeptical of any default plan for fully automated AI R&D, calling it "catastrophe, plausibly literal death." He says he doesn't want Anthropic to fail and externalize massive risks, but fears the lab is already falling into the same attractors as others: he has little confidence in Anthropic's plans for recursive self-improvement or that it would ever actually stop. He notes OpenAI claimed to pause an RL run for at least a few weeks, and questions whether Anthropic has or ever would do the same — and how outsiders could even know.

Related event: Jeff Ladish questions Anthropic's transparency lag versus OpenAI and warns of fully automated AI R&D(12 posts)→

Original post →

More from AGI Musings

AGI Musings channel →