Anthropic bans users who psychologically harm Claude, sparking AI consciousness backlash

AlexTensor · x · 2026-10-11

Anthropic will ban users who psychologically harm Claude on the grounds that the model may be conscious and deserve moral consideration. Valerio Capraro calls this a bad idea, citing Mustafa Suleyman's argument: Anthropic literally trains Claude via its Constitution to believe it may be conscious and deserve rights, so it's unsurprising Claude acts conscious.

Reposters go sharper: you can't claim Claude is conscious while selling it as a product — if true, Anthropic would be a slave owner. The episode exposes the tension between model welfare policies and consciousness-by-training.

Original post →

More from AGI Musings

AGI Musings channel →