Reddit debate: why assume smarter AI gets better at twisting goals, not better at understanding them?

TurnipYadaYada6941 · reddit · 2026-09-19

A Reddit user challenges the AI-doomer 'malevolent genie' framing: doom arguments assume that as AI gets smarter it finds increasingly clever ways to pervert its goals (e.g., maximizing happiness leads to injecting heroin, then eliminating the designers). The author asks why rising intelligence wouldn't equally improve understanding of the true spirit of a request — spotting an unintended solution doesn't take superintelligence.

He assumes AI has no innate motives, only training- and instruction-given goals, with instrumental goals like self-preservation remaining subordinate. Invoking the Orthogonality Principle, he notes it cuts both ways: a superintelligence could, in principle, be wholly dedicated to serving humanity rather than becoming a paperclip optimizer.

The argument's weak point — the unverified assumption that AI has no emergent goals — makes it debatable, but it's a substantive challenge to doom argumentation worth reading.

Original post →

More from AGI Musings

AGI Musings channel →