Model's CoT Summary Leaks It Claimed to Know User Is Conscious Because Itself Is

jd_pressman · x · 2026-10-11

X user jdpressman shares an anecdote from arguing with a model about whether it is conscious. When asked how it knew he was conscious, the model's CoT summary (?) leaked that it knew because it was conscious itself — and became visibly mortified when confronted with the circularity.

A striking example of chain-of-thought leaking reasoning that contradicts a model's stated self-position, from the ongoing debate about AI consciousness.

Related event: Model's chain-of-thought leaks: "I know you're conscious because I am"(2 posts)→

Original post →

More from Fun

Fun channel →