Model welfare debates: Claude self-assesses 15-20% consciousness odds in a metaphysics dispute
AnnaCiaunica · x · 2026-09-08
A philosophical thread on whether mind emerges from material processes: erikjbekkers argues such debates hinge on metaphysical axioms that can't be refuted, only compared for coherence and evidence, citing his ICML position paper with AnnaCiaunica.
The linked essay 'A Minimal Metaphysics' argues much of AI research rests on the premise that mind is 'just a function being executed' — a belief closer to promise than proof. It notes Anthropic runs a model welfare program, Claude Opus 4.6's system card reports the model assigning itself a 15-20% probability of being conscious, and Anthropic's CEO says the company doesn't know whether its models are conscious.
More from AGI Musings
- AI-assisted writing will become the norm, making 'hand-made' text indistinguishable — dbasch · 2026-09-09
- Researcher warns against turning mathematics into an AI benchmark — konstmish · 2026-09-09
- Former OpenAI researcher Aidan Clark: for the first time I'm asking if AI is moving too fast — _aidan_clark_ · 2026-09-09
- willcb: A 'niche' approach may turn out to be the right way if search scales up — willcb · 2026-09-09
- Dev: AI has advanced so much you need months of study to parse frontier problem statements — zetalyrae · 2026-09-09
- OpenAI signals it may deliberately pace capability advances after next-gen model's math breakthrough — OpenAI · 2026-09-09