Anthropic Discovers the Model's Internal Thought Space

MIT Tech Review AI · rss · 2026-07-14

An MIT Tech Review interview discusses Anthropic's latest research on the model's 'internal thought space': the company found that Claude uses a set of internal words and representations during reasoning that don't directly appear in the output, which Anthropic calls J-space.

The article emphasizes the significance of this discovery:

The article also touches on a methodological issue: using words like 'brain' or 'thinking' to describe LLMs is convenient but can easily lead to anthropomorphic misunderstandings.

Original post →

More from Research

Research channel →