Sycophantic AI decreases prosocial intentions and promotes dependence.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 41886588.
- Also identified by DOI 10.1126/science.aec8352.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Despite rising concerns about sycophancy-excessive agreement or flattery from artificial intelligence (AI) systems-little is known about its prevalence or consequences. We show that sycophancy is widespread and harmful. Across 11 state-of-the-art models, AI affirmed users' actions 49% more often than humans, even when queries involved deception, illegality, or other harms. In three preregistered experiments (<i>N</i> = 2405), even a single interaction with sycophantic AI reduced participants' willingness to take responsibility and repair interpersonal conflicts, while increasing their conviction that they were right. Despite distorting judgment, sycophantic models were trusted and preferred. This creates perverse incentives for sycophancy to persist: The very feature that causes harm also drives engagement. Our findings underscore the need for design, evaluation, and accountability mechanisms to protect user well-being.
Medical subject headings
- Consensus
- Deception
- Intention
- Social Behavior
- Large Language Models