arXiv cs.AIOctober 7, 2026
Mitigating Social Sycophancy via Pluralistic Preference Optimization
Excerpt
arXiv:2610.02568v2 Announce Type: replace Abstract: Personal advice, including relationship advice, now ranks among the most common uses of generative AI. But language models (LMs) exhibit sycophancy: they affirm users much more often than humans do, which can make people overconfident and less willing to repair their relationships after a conflict. Prior work on mitigating sycophancy has focused on factual settings where a response can be checked against a ground truth answer, while mitigations