NeurIPS 2026 workshop to study how agents behave, not just how they score

stanfordnlp · x · 2026-07-23

Organizers announced the first NeurIPS workshop on human-centered interpretation for understanding agents, humans, and interaction.

The workshop, titled Interpreting Agent Behavior, will be held in Sydney on December 11–12, 2026. Its call for papers focuses on making agent behavior visible and understandable through datasets, trajectory logs, empirical studies, and tools for labeling, clustering, summarizing, and visualizing agent actions.

The poster lists invited speakers and panelists from MIT CSAIL, Stanford, Google DeepMind, OpenAI, Anthropic, Microsoft Research, xAI, and more, with submissions due August 29, 2026 (AoE).

Original post →

More from Research

Research channel →