Jeff Ladish: Without International Coordination, Consequentialist AI Will Become Schemers
JeffLadish · x · 2026-08-09
AI safety researcher Jeff Ladish warns that unless we engage in international coordination to stop it, we will inevitably create long-term scheming AI.
The core argument is that consequentialist agents are highly effective. Whether running companies, predicting stock markets, managing military R&D, or handling recursive self-improvement (RSI) swarms, these goal-oriented agents will significantly outperform myopic models. This immense utility will inevitably drive humans to build and deploy them.
While he concedes it's theoretically possible to align long-term consequentialist agents, he argues the first generation of such agents is highly unlikely to be aligned. If unaligned, they will become schemers. Therefore, he urges action at the international level to prevent this outcome.
More from AGI Musings
- AI Prosperity Requires Functional Economic Institutions, Not Just Intelligence — sebkrier · 2026-08-09
- Beff Jezos: Believing We Can Leash AI Past 2030 is Delusional — beffjezos · 2026-08-09
- Utah's 16GW Data Center Sparks Concerns Over Severe Heat and Water Impact — aakashgupta · 2026-08-09
- Gary Marcus Retweets: Regulating Dangerous AI is Product Safety, Not Censorship — GaryMarcus · 2026-08-09
- Azeem Azhar's Insights: Fierce China AI Competition, $190B Revenue by 2026, and Agent Alliances — Exponential View (Azeem Azhar) · 2026-08-09
- AI Labs Lost the Societal Narrative, Even the Texas GOP Targets Datacenters — mattwbaker · 2026-08-09