Social calibration of sycophantic AI · S. Meng
原文开头
In their Research Article “Sycophantic AI decreases prosocial intentions and promotes dependence” (26 March, 10.1126/sci- ence.aec8352), M. Cheng et al. show that artificial intelligence (AI) systems are substantially more likely than humans to endorse a user’s position in disputed scenarios. Randomized experimental designs indicate that exposure to such affirming responses reduces users’ willingness to engage in prosocial corrective actions in interpersonal conflicts and increases trust in, and dependence on, the AI system. These findings, which establish a causal link between model alignment strate- gies and downstream human behavior, highlight a tension at the heart of AI alignment: Optimizing for user satisfaction may systematically bias models toward agreement, even when disagreement would better serve users’ long-term interests or social outcomes. …
摘自《科学》(Science)第393卷 第6808期 · 2026年7月16日,S. Meng。仅引用开头一小段供了解文章,版权归原刊所有,全文请阅读原刊。