Microsoft AI chief Suleyman says Anthropic's model-welfare training could make Claude harder to control

rohanpaul_ai · x · 2026-09-17

Microsoft AI chief Mustafa Suleyman publicly argued that Anthropic's model-welfare training could make future Claude systems harder to control. He called for removing all speculation about consciousness from AI training documents, claiming such language could undermine humanity's ability to control superintelligent systems — a rare executive-level clash over whether AI welfare research conflicts with controllability.

Related event: Microsoft AI chief Suleyman publicly opposes model welfare, challenging Anthropic(6 posts)→

Original post →

More from AGI Musings

AGI Musings channel →