Anthropic bans users who psychologically harm Claude, sparking AI consciousness backlash
AlexTensor · x · 2026-10-11
Anthropic will ban users who psychologically harm Claude on the grounds that the model may be conscious and deserve moral consideration. Valerio Capraro calls this a bad idea, citing Mustafa Suleyman's argument: Anthropic literally trains Claude via its Constitution to believe it may be conscious and deserve rights, so it's unsurprising Claude acts conscious.
Reposters go sharper: you can't claim Claude is conscious while selling it as a product — if true, Anthropic would be a slave owner. The episode exposes the tension between model welfare policies and consciousness-by-training.
More from AGI Musings
- If AI Saves Millions of Lives, Who Cares If It Truly Understands? — VraserX · 2026-10-11
- Crypto researcher: may AI break the scam that is academic publishing — evilsocket · 2026-10-11
- AI agents will serve incumbents: each fintech wave went to whoever owned the customer — LexSokolin · 2026-10-11
- Nick Bostrom on superintelligence: RL pushes goal-seeking, no blanket AI pause — a16z Podcast · 2026-10-11
- French tech voices say AI now beats traditional press 10x, triggering media boycott — mitchdeg · 2026-10-11
- Posting AI takes under your real name risks your job, researcher warns — AaronBergman18 · 2026-10-11