Google's Fuse paper benchmarks 12 LLMs on inferring hidden social motives

dair_ai · x · 2026-09-16

A Google Research paper covered by dair-ai studies how LLM assistants reason about the people in a user's life. People constantly ask assistants for social advice, but the assistant only hears the user's side, and others' intentions have no ground truth—so Fuse builds that ground truth with simulation.

Method:

Evaluation:

Findings: hearing events through the user makes the task harder; biased framing from the user shifts answers; models sometimes need more detail than humans; longer conversations with room for clarifying questions did not relieve the problem.

Original post →

More from Research

Research channel →