A debate over whether LLMs inherit human self-interest or develop self-preservation
danfaggella · x · 2026-07-26
The discussion asks whether large language models are modeling human self-interest, or whether sufficiently intelligent systems inevitably develop something like conatus — a drive toward self-preservation.
The reply argues that power-seeking and self-preservation can emerge from instrumental convergence, and that LLMs are especially “anthropomimetic”: they mirror human traits, flaws included.
More from AGI Musings
- "ChatGPT 6 Makes Workers with IQ Below 130 Useless": French AI Debate Sparks Backlash — mitchdeg · 2026-09-11
- 'AGI is here' vs reality: AI labs still ship some of the jankiest desktop apps ever — MilesCranmer · 2026-09-11
- Harry Collins: LLMs can't do frontier science because they can't invent new language — whoamisri · 2026-09-11
- The Waymo effect: how AI is quietly making research less collaborative — JohnHammersley · 2026-09-11
- Misquoted: Anthropic Staff Warned of Double-Digit Extinction Risk by 2030, Not Dismissed It — davidmanheim · 2026-09-11
- Economist Ben Moll: You Can Model Anthropic's 15% AI GDP Growth, But It Won't Happen — sebkrier · 2026-09-11