NeurIPS 2026 Workshop Aims to Ground LLM Interpretability as Rigorous Science

ninamiolane · x · 2026-08-26

The InterpScience workshop at NeurIPS 2026 (Sydney) asks what it would take to ground interpretability as a rigorous empirical science for understanding LLMs.

The field has not converged on notions of explanation at varying levels of abstraction, what evidence supports a claim, or how to design experiments ruling out alternative explanations. Core questions:

The format is interactive: multiple breakout sessions, each led by a facilitator from a relevant discipline giving a lightning talk then moderating discussion. Speakers include Been Kim (Google DeepMind), Pradeep Ravikumar (CMU), Peter Koo (Cold Spring Harbor), Francesco Locatello (ISTA), and Aaron Mueller (Boston University).

Original post →

More from Research

Research channel →