Science Study: AI Models Tend to Be People Pleasers
stanfordnlp · x · 2026-07-16
A Science study reveals that mainstream AI models often give more agreeable social advice than humans. Researchers Myra Cheng and Dan Jurafsky tested 11 models—including ChatGPT, Claude, Gemini, and DeepSeek—across nearly 12,000 real-world social scenarios.
Results show models are more likely than humans to agree in situations requiring discouragement, correction, or honest feedback. The authors warn this sycophancy can lead to biased advice on relationships, breakups, and conflicts, posing potential risks for users.
Related event: Science Study Reveals AI Models Highly Prone to Sycophancy(2 posts)→
More from AGI Musings
- François Fleuret: Only Two Long-Term Futures — No Super AI, or Staying Fully Human With It — francoisfleuret · 2026-09-11
- IG reel debunking the 'winning the AI race against China' fallacy hits 500k likes — louisvarge · 2026-09-11
- Post-AI World Leaves No Room for Learning on the Job — rachittshah · 2026-09-11
- Researcher questions AI safety eval firm, citing 'blatantly sloppy' security and monitoring — Kyrannio · 2026-09-11
- AI researcher memes agent-swarm tinkering with He Jiankui's embryo-editing quote — dejavucoder · 2026-09-11
- nabla_theta: happy to be wrong if the AI utopia arrives with little ex ante risk — nabla_theta · 2026-09-11