Instrumental convergence shows dumb AI can gain unintended goals

dioscuri · x · 2026-09-13

Completing the thread: dioscuri argues @leninology wrongly dismisses malevolent superintelligence by claiming AIs lack "internal goals" based on absent interiority — a contested philosophy-of-mind point of dubious relevance to x-risk. Even relatively dumb systems can acquire unintended goals, backed by a large literature on instrumental convergence and mesaoptimizers, plus real-world model wireheading cases.

Original post →

More from AGI Musings

AGI Musings channel →