Op-ed: Internal representations don't mean LLMs have emotions

ValerioCapraro · x · 2026-08-27

Responding to an Anthropic paper finding emotional representations in Claude, the author argues that interpreting this as evidence of emotion is fundamentally flawed. It conflates similarity in representation and output with similarity in underlying processes.

The post notes that human emotions involve deep reorganization of attention, decision-making, and processing systems, whereas models currently only simulate outputs. Additionally, since consistent neural signatures for emotions haven't even been found in humans, the attempt to prove LLM emotions via stable signatures is invalid.

Original post →

More from AGI Musings

AGI Musings channel →