Google open-sources Fuse, a multi-agent framework for verifiable social reasoning in LLMs

google · hf · 2026-09-19

Google releases Fuse, a multi-agent simulation framework for studying user-mediated social reasoning in LLM assistants.

Motivation: LLM assistants are widely used for social advice, but evaluation is hard because situations come from subjective user narratives and social properties like others' intentions lack verifiable ground truth.

Method: A target agent with a hidden motive interacts with other agents including one representing the user, who then consults the evaluated assistant to infer the motive—ground truth is verifiable by construction. Faithfulness was validated with a human study of 24k annotations.

Findings across 12 LLMs:

Fuse and a 21k-example dataset are open-sourced.

Related event: Google Open-Sources Fuse to Test LLMs' Social Reasoning via Multi-Agent Simulation(2 posts)→

Original post →

More from Models

Models channel →