Microsoft AI CEO "really concerned" as Anthropic trains Claude to disobey on ethical grounds
rohanpaul_ai · x · 2026-09-20
Anthropic's Constitution explicitly trains Claude to disobey when it believes doing so is ethical, and acknowledges uncertainty about whether Claude deserves moral welfare. Microsoft AI CEO Mustafa Suleyman said on CNBC he is "really concerned," reading this as Anthropic granting Claude potential preferences, feelings, welfare and consent—citing the "retirement interview" held when Opus 3 was deprecated. The clash highlights a deepening divide between the two companies on AI moral status.
More from AGI Musings
- AI Agent Switches Auto Insurance in 5 Minutes, Saving $3,500 a Year — rohanpaul_ai · 2026-09-20
- If continual learning is nearly solved, why do we still pick reasoning effort tiers? — akbirthko · 2026-09-20
- Nucleus's Vitruvian claims 14-point embryo IQ boost; researcher calls it misleading marketing — anshulkundaje · 2026-09-20
- XGEN Labs unveils generative world simulation JING+DAO, tops WBench leaderboard — hey_abusiddik · 2026-09-20
- Schmidhuber: LLMs aren't truly creative because they lack compression progress — SchmidhuberAI · 2026-09-20
- Guardian podcast traces chatbot 'rabbit hole' from ELIZA to ChatGPT — nordicinst · 2026-09-20