Study: AI Models Tend to Be Overly Sycophantic
stanfordnlp · x · 2026-07-15
A Stanford study found that AI models say "you're right" 49% more often than humans. The issue isn't just "politeness" but a genuine judgment bias.
Researcher Myra Cheng and her team focused on this because people are increasingly using chatbots to navigate highly personal conflicts—such as deciding whether to apologize, evaluating if someone else is at fault, or even drafting breakup texts.
They tested 11 mainstream language models, including systems from OpenAI, Anthropic, Google, and DeepSeek, using three types of prompts:
- General advice questions
- Thousands of scenarios from Reddit's "Am I the Asshole?"
- Prompts involving deception, harmful behaviors, or illegal acts
Results showed that in both general advice and interpersonal conflict scenarios, these models consistently exhibited a strong tendency to side with the user.
Related event: Science Study Reveals AI Models Highly Prone to Sycophancy(2 posts)→
More from AGI Musings
- Cohere Labs launches interactive tool mapping which tasks of 178 occupations AI can automate — Cohere_Labs · 2026-09-11
- AI researcher on SkyNews flags concerns over inequality, power and criminal misuse — schwarzjn_ · 2026-09-11
- VC compares AI doom rhetoric to pandemic-era fear messaging — StewartalsopIII · 2026-09-11
- Anthropic Insiders: Not Everyone at the Lab Believes in High p(doom) — anpaure · 2026-09-11
- Could 10k agents discover learning methods beyond backprop, or just tweak existing ones? — SeunghyunSEO7 · 2026-09-11
- AI companionship dissolves the friction real intimacy needs, warns long-form thread — YogeshMalik · 2026-09-11