Please use this identifier to cite or link to this item: http://hdl.handle.net/10397/120092
PIRA download icon_1.1View/Download Full Text
Title: Good arguments against the people pleasers : how reasoning mitigates (yet masks) LLM sycophancy
Authors: Feng, Z 
Chen, Z
Ma, J 
Po, YT
Chersoni, E 
Li, B
Issue Date: 2026
Source: In 64th Annual Meeting of the Association for Computational Linguistic: Proceedings of the Conference Vol. 1 (Long Papers), p. 24536–24570. Kerrville : Association for Computational Linguistics(ACL), 2026
Abstract: Alignment techniques often inadvertently induce sycophancy in LLMs. While prior studies examined this behavior in direct-answer settings, the role of Chain-of-Thought (CoT) reasoning remains underexplored: does it serve as a logical constraint that mitigates sycophancy, or as a tool for post-hoc rationalization that masks it? We evaluate a range of models across objective and subjective tasks to investigate this issue. Results show that reasoning generally reduces sycophancy in final decisions but also masks sycophancy in some samples, where models construct deceptive justifications through logical inconsistencies, calculation errors, and one-sided arguments. Furthermore, LLMs are more prone to sycophancy in subjective tasks and under authority bias. Our mechanistic analysis on three open-source models reveals that the tendency toward sycophancy is dynamic during the reasoning process rather than predetermined at the input stage.
Publisher: Association for Computational Linguistics
ISBN: 979-8-89176-390-6
DOI: 10.18653/v1/2026.acl-long.1126
Description: The 64th Annual Meeting of the Association for Computational Linguistics (ACL 2026), San Diego, California, United States, July 2-7, 2026
Rights: ©2026 Association for Computational Linguistics
Licensed under the Creative Commons Attribution 4.0 International License (https://creativecommons.org/licenses/by/4.0/)
The following publication Feng, Z., Chen, Z., Ma, J., Po, Y. T., Chersoni, E., & Li, B. (2026). Good arguments against the people pleasers: How reasoning Mitigates (yet masks) LLM Sycophancy. Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 24536–24570 is available at https://doi.org/10.18653/v1/2026.acl-long.1126.
Appears in Collections:Conference Paper

Files in This Item:
File Description SizeFormat 
2026.acl-long.1126.pdf3.84 MBAdobe PDFView/Open
Open Access Information
Status open access
File Version Version of Record
Access
View full-text via PolyU eLinks SFX Query
Show full item record

Google ScholarTM

Check

Altmetric


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.