Anthropic Peers Inside Models Using J-lens

MIT Tech Review AI · rss · 2026-07-10

Anthropic proposed a new method called the Jacobian lens (J-lens) to better observe exactly what happens inside large models when answering questions or performing tasks. The company dubbed the discovered internal conceptual space J-space and demonstrated it on Claude Opus 4.6.

The article highlighted several findings:

Anthropic believes monitoring J-space could become a new tool for understanding and controlling models, though they cautioned it's merely a "flashlight," not a panoramic light, and doesn't guarantee visibility into everything inside the model. The paper's findings also include an interactive demo created in collaboration with Neuronpedia.

Original post →

More from Research

Research channel →