DeepMind Discusses Model Interpretability

GoogleDeepMind · x · 2026-07-11

A Google DeepMind podcast invited @NeelNanda5 to discuss interpretability research, focusing on "reverse engineering" how neural networks learn and think.

Key points include:

Related event: DeepMind Discusses Chain of Thought and Mechanistic Interpretability(4 posts)→

Original post →

More from Research

Research channel →