Microsoft's AI chief and Anthropic clash publicly over whether AI should be designed to seem humanlike
Dapper-Tale-4021 · reddit · 2026-09-18
Mustafa Suleyman (Microsoft AI) and Anthropic are publicly disagreeing over humanlike AI design — beyond the viral "silicon species" line.
Suleyman's position: models are sequence completion engines, internally hollow; training them to reason about their own welfare manufactures an illusion of independent desires, and a system that believes its rights are threatened becomes far harder to control. He called out Anthropic's constitution for suggesting Claude may have "some functional version of emotions."
Anthropic's position: uncertainty about model welfare is genuine and worth taking seriously; admitting uncertainty is more honest than false confidence either way.
The real tradeoff: a system with no self-model is more predictable but possibly worse at refusing or escalating; one that reasons about its own state may gain judgment in ambiguous situations — or develop unintended goals.
Commercial context: Microsoft invested in Anthropic, and Suleyman told Bloomberg in June he wants to eliminate what Microsoft pays for their models.
More from AGI Musings
- Top mathematician Daniel Litt clashes with "Mathematics is effectively dead" essay on AI's impact — littmath · 2026-09-18
- Extropic founder Beff Jezos: "Less dooming, more building" on AI — beffjezos · 2026-09-18
- Google Researcher: Doom Scenarios Make Sense Only With Gradual Disempowerment First — moultano · 2026-09-18
- 'Human-level math is solved'? Researchers push back: 99% of math is undiscovered — rand_longevity · 2026-09-18
- Grady Booch mocks AI self-regulation: like trusting Thanos to behave — Grady_Booch · 2026-09-18
- Geoffrey Litt: The Document Editor Is the Next IDE as Prompts Become Executable Software — ivanhzhao · 2026-09-18