PlurPO: training LLMs to curb social sycophancy that discourages relationship repair

RishiBommasani · x · 2026-10-06

A thread by HatgisKessell: people increasingly turn to AI for personal advice, but LLMs endorse users far more often than humans do. The consequence: people become less willing to repair relationships after conflicts. The team proposes PlurPO, a method to build a pluralistic preference dataset for post-training models to mitigate this social sycophancy.

Original post →

More from Research

Research channel →