Study finds most accurate models are not always the most preferred by users

soumitrashukla9 · x · 2026-08-27

A highlight of work led by Mina Lee finds that the most accurate models are not always the ones human users prefer the most.

For instance, Opus and Sonnet are great at performing tasks on their own but not in providing assistance, while GPT-5-Mini was strong in both dimensions. Gemini models were found to be stronger as assistants than as automators.

Related event: CentaurBench: The Strongest Models Aren't Always the Best Assistants(3 posts)→

Original post →

More from Research

Research channel →