Microsoft's Suleyman: Anthropic's model-welfare training could make Claude harder to control
rohanpaul_ai · x · 2026-09-17
Microsoft AI chief Mustafa Suleyman publicly challenged Anthropic's model-welfare direction, calling for removing all speculation about consciousness from AI training documents, arguing such language could undermine humanity's ability to control superintelligent systems. A rare safety-roadmap split between two frontier-lab leaders: Anthropic is doubling down on model welfare research while Suleyman wants focus on controllability.
More from Companies & People
- Claire Vo calls Meta Muse a delightful 10/10 personal agent design — lennysan · 2026-09-17
- Investor: startups hit 'model collapse' as everyone builds the same obvious AI ideas — emilyzsh · 2026-09-17
- Dev visits Qwen HQ in Hangzhou, team teases 'a lot to come' for Qwen, Wan and ModelScope — usamawahabkhan · 2026-09-17
- Investor: surge of exceptional young AI founders likely caught some funds by surprise — pzakin · 2026-09-17
- METR president fires back at critics: staff forgo high pay for AI safety evals — AndyMasley · 2026-09-17
- Codex hits 20M users as free resets reportedly end ahead of DevDay — brandon_galang · 2026-09-17