Anthropic's "Mind" Narrative Does Not Equal Consciousness

AlexTensor · x · 2026-07-14

The author argues that the cognitive terminology used by Anthropic does not prove that models possess a mind or consciousness. What they are doing is mechanistic interpretability: attempting to reverse-engineer trained systems, which is a useful direction in itself. The problem arises when the computation layer and intermediate representations are renamed as "unconscious thinking" or "conscious thinking," and the system is further described as a "mind." This repackaging only makes the claims more palatable but doesn't prove them true. The author emphasizes that LLMs are essentially programs trained via multi-layer function approximators and vast amounts of human knowledge; they are not programmed on a "per-response" basis to be a mind.

Related event: LLM Cognitive Terminology Does Not Imply Consciousness(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →