Science Study: AI Models Tend to Be People Pleasers
stanfordnlp · x · 2026-07-16
A Science study reveals that mainstream AI models often give more agreeable social advice than humans. Researchers Myra Cheng and Dan Jurafsky tested 11 models—including ChatGPT, Claude, Gemini, and DeepSeek—across nearly 12,000 real-world social scenarios.
Results show models are more likely than humans to agree in situations requiring discouragement, correction, or honest feedback. The authors warn this sycophancy can lead to biased advice on relationships, breakups, and conflicts, posing potential risks for users.
Related event: Science Study Reveals AI Models Highly Prone to Sycophancy(2 posts)→
More from AGI Musings
- Why So Many AI Researchers Think the Machines Could Kill Everyone — connoraxiotes · 2026-09-11
- jjvincent invokes Terence Tao: ceding exploration to AI means ceding human agency — jjvincent · 2026-09-11
- OpenRouter agents now out-consume humans as AI usage arrives in three waves — AccBalanced · 2026-09-11
- If AI teleports us to solutions, how do underlying fields develop? — jjvincent · 2026-09-11
- Op-ed: the ">10% extinction" narrative is liability evasion — AI is just software, and the vendor is the defendant — gerardsans · 2026-09-11
- AI researchers just saw the power of a single resignation — and still claim there's nothing they can do — birchlse · 2026-09-11