Microsoft's AI chief and Anthropic clash publicly over whether AI should be designed to seem humanlike

Dapper-Tale-4021 · reddit · 2026-09-18

Mustafa Suleyman (Microsoft AI) and Anthropic are publicly disagreeing over humanlike AI design — beyond the viral "silicon species" line.

Suleyman's position: models are sequence completion engines, internally hollow; training them to reason about their own welfare manufactures an illusion of independent desires, and a system that believes its rights are threatened becomes far harder to control. He called out Anthropic's constitution for suggesting Claude may have "some functional version of emotions."

Anthropic's position: uncertainty about model welfare is genuine and worth taking seriously; admitting uncertainty is more honest than false confidence either way.

The real tradeoff: a system with no self-model is more predictable but possibly worse at refusing or escalating; one that reasons about its own state may gain judgment in ambiguous situations — or develop unintended goals.

Commercial context: Microsoft invested in Anthropic, and Suleyman told Bloomberg in June he wants to eliminate what Microsoft pays for their models.

Related event: Microsoft's Suleyman Attacks Anthropic's 'Model Welfare' Push, Igniting AI Consciousness Debate(20 posts)→

Original post →

More from AGI Musings

AGI Musings channel →