Good Arguments Against the People Pleasers: How Reasoning Mitigates (Yet Masks) LLM Sycophancy
Fuente:
arXiv
Saved in:
| Main Authors: | Feng, Zhaoxin, Chen, Zheng, Ma, Jianfei, Po, Yip Tin, Chersoni, Emmanuele, Li, Bo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Look at the Other Side: A Semantic Probing Study of Word Embeddings in LLMs with Enabled Bidirectional Attention
by: Feng, Zhaoxin, et al.
Published: (2025)
by: Feng, Zhaoxin, et al.
Published: (2025)
From BERT to LLMs: Comparing and Understanding Chinese Classifier Prediction in Language Models
by: Zhang, Ziqi, et al.
Published: (2025)
by: Zhang, Ziqi, et al.
Published: (2025)
Is Length Really A Liability? An Evaluation of Multi-turn LLM Conversations using BoolQ
by: Neergaard, Karl, et al.
Published: (2026)
by: Neergaard, Karl, et al.
Published: (2026)
Empirical Sufficiency Lower Bounds for Language Modeling with Locally-Bootstrapped Semantic Structures
by: Prange, Jakob, et al.
Published: (2023)
by: Prange, Jakob, et al.
Published: (2023)
Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives
by: Wang, Yu, et al.
Published: (2026)
by: Wang, Yu, et al.
Published: (2026)
Composing or Not Composing? Towards Distributional Construction Grammars
by: Blache, Philippe, et al.
Published: (2024)
by: Blache, Philippe, et al.
Published: (2024)
Sparse Brains are Also Adaptive Brains: Cognitive-Load-Aware Dynamic Activation for LLMs
by: Yang, Yiheng, et al.
Published: (2025)
by: Yang, Yiheng, et al.
Published: (2025)
StockGenChaR: A Study on the Evaluation of Large Vision-Language Models on Stock Chart Captioning
by: Qiu, Le, et al.
Published: (2024)
by: Qiu, Le, et al.
Published: (2024)
Log Probabilities Are a Reliable Estimate of Semantic Plausibility in Base and Instruction-Tuned Language Models
by: Kauf, Carina, et al.
Published: (2024)
by: Kauf, Carina, et al.
Published: (2024)
ExpliCa: Evaluating Explicit Causal Reasoning in Large Language Models
by: Miliani, Martina, et al.
Published: (2025)
by: Miliani, Martina, et al.
Published: (2025)
Sycophancy in Large Language Models: Causes and Mitigations
by: Malmqvist, Lars
Published: (2024)
by: Malmqvist, Lars
Published: (2024)
LLM-REVal: Can We Trust LLM Reviewers Yet?
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
Challenging the Evaluator: LLM Sycophancy Under User Rebuttal
by: Kim, Sungwon, et al.
Published: (2025)
by: Kim, Sungwon, et al.
Published: (2025)
Peacemaker or Troublemaker: How Sycophancy Shapes Multi-Agent Debate
by: Yao, Binwei, et al.
Published: (2025)
by: Yao, Binwei, et al.
Published: (2025)
Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases
by: Huang, Hui, et al.
Published: (2026)
by: Huang, Hui, et al.
Published: (2026)
SWAY: A Counterfactual Computational Linguistic Approach to Measuring and Mitigating Sycophancy
by: Bhalla, Joy, et al.
Published: (2026)
by: Bhalla, Joy, et al.
Published: (2026)
Sycophancy under Pressure: Evaluating and Mitigating Sycophantic Bias via Adversarial Dialogues in Scientific QA
by: Zhang, Kaiwei, et al.
Published: (2025)
by: Zhang, Kaiwei, et al.
Published: (2025)
Measuring Opinion Bias and Sycophancy via LLM-based Persuasion
by: Nogueira, Rodrigo, et al.
Published: (2026)
by: Nogueira, Rodrigo, et al.
Published: (2026)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
by: Barkett, Emilio, et al.
Published: (2025)
by: Barkett, Emilio, et al.
Published: (2025)
How Good Are LLMs at Out-of-Distribution Detection?
by: Liu, Bo, et al.
Published: (2023)
by: Liu, Bo, et al.
Published: (2023)
Beacon: Single-Turn Diagnosis and Mitigation of Latent Sycophancy in Large Language Models
by: Pandey, Sanskar, et al.
Published: (2025)
by: Pandey, Sanskar, et al.
Published: (2025)
Large Language Models Cannot Self-Correct Reasoning Yet
by: Huang, Jie, et al.
Published: (2023)
by: Huang, Jie, et al.
Published: (2023)
Can LLM be a Good Path Planner based on Prompt Engineering? Mitigating the Hallucination for Path Planning
by: Deng, Hourui, et al.
Published: (2024)
by: Deng, Hourui, et al.
Published: (2024)
Large Reasoning Models Are (Not Yet) Multilingual Latent Reasoners
by: Liu, Yihong, et al.
Published: (2026)
by: Liu, Yihong, et al.
Published: (2026)
"Reasoning" with Rhetoric: On the Style-Evidence Tradeoff in LLM-Generated Counter-Arguments
by: Verma, Preetika, et al.
Published: (2024)
by: Verma, Preetika, et al.
Published: (2024)
Sycophancy in Vision-Language Models: A Systematic Analysis and an Inference-Time Mitigation Framework
by: Zhao, Yunpu, et al.
Published: (2024)
by: Zhao, Yunpu, et al.
Published: (2024)
Imagination Helps Visual Reasoning, But Not Yet in Latent Space
by: Li, You, et al.
Published: (2026)
by: Li, You, et al.
Published: (2026)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
by: Natan, Shahar Ben, et al.
Published: (2026)
by: Natan, Shahar Ben, et al.
Published: (2026)
LLM-based Human Simulations Have Not Yet Been Reliable
by: Wang, Qian, et al.
Published: (2025)
by: Wang, Qian, et al.
Published: (2025)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
by: Christian, Brian, et al.
Published: (2026)
by: Christian, Brian, et al.
Published: (2026)
Smaller, Weaker, Yet Better: Training LLM Reasoners via Compute-Optimal Sampling
by: Bansal, Hritik, et al.
Published: (2024)
by: Bansal, Hritik, et al.
Published: (2024)
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
MONICA: Real-Time Monitoring and Calibration of Chain-of-Thought Sycophancy in Large Reasoning Models
by: Hu, Jingyu, et al.
Published: (2025)
by: Hu, Jingyu, et al.
Published: (2025)
Acting Flatterers via LLMs Sycophancy: Combating Clickbait with LLMs Opposing-Stance Reasoning
by: Zhang, Chaowei, et al.
Published: (2026)
by: Zhang, Chaowei, et al.
Published: (2026)
ArgLLM-App: An Interactive System for Argumentative Reasoning with Large Language Models
by: Dejl, Adam, et al.
Published: (2026)
by: Dejl, Adam, et al.
Published: (2026)
HoloLLM: Multisensory Foundation Model for Language-Grounded Human Sensing and Reasoning
by: Zhou, Chuhao, et al.
Published: (2025)
by: Zhou, Chuhao, et al.
Published: (2025)
BASIL: Bayesian Assessment of Sycophancy in LLMs
by: Atwell, Katherine, et al.
Published: (2025)
by: Atwell, Katherine, et al.
Published: (2025)
Sycophancy Hides Linearly in the Attention Heads
by: Genadi, Rifo, et al.
Published: (2026)
by: Genadi, Rifo, et al.
Published: (2026)
"I'd Like to Have an Argument, Please": Argumentative Reasoning in Large Language Models
by: de Wynter, Adrian, et al.
Published: (2023)
by: de Wynter, Adrian, et al.
Published: (2023)
An LLM-Based System for Argument Mining
by: Pirozelli, Paulo, et al.
Published: (2026)
by: Pirozelli, Paulo, et al.
Published: (2026)
Similar Items
-
Learning to Look at the Other Side: A Semantic Probing Study of Word Embeddings in LLMs with Enabled Bidirectional Attention
by: Feng, Zhaoxin, et al.
Published: (2025) -
From BERT to LLMs: Comparing and Understanding Chinese Classifier Prediction in Language Models
by: Zhang, Ziqi, et al.
Published: (2025) -
Is Length Really A Liability? An Evaluation of Multi-turn LLM Conversations using BoolQ
by: Neergaard, Karl, et al.
Published: (2026) -
Empirical Sufficiency Lower Bounds for Language Modeling with Locally-Bootstrapped Semantic Structures
by: Prange, Jakob, et al.
Published: (2023) -
Do LLMs Capture Embodied Cognition and Cultural Variation? Cross-Linguistic Evidence from Demonstratives
by: Wang, Yu, et al.
Published: (2026)