Sycophancy Claims about Language Models: The Missing Human-in-the-Loop
Fuente:
arXiv
Saved in:
| Main Authors: | Batzner, Jan, Stocker, Volker, Schmid, Stefan, Kasneci, Gjergji |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GermanPartiesQA: Benchmarking Commercial Large Language Models and AI Companions for Political Alignment and Sycophancy
by: Batzner, Jan, et al.
Published: (2024)
by: Batzner, Jan, et al.
Published: (2024)
Whose Personae? Synthetic Persona Experiments in LLM Research and Pathways to Transparency
by: Batzner, Jan, et al.
Published: (2025)
by: Batzner, Jan, et al.
Published: (2025)
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI
by: Schirmer, Miriam, et al.
Published: (2024)
by: Schirmer, Miriam, et al.
Published: (2024)
Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks
by: Kasneci, Enkelejda, et al.
Published: (2026)
by: Kasneci, Enkelejda, et al.
Published: (2026)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
by: Natan, Shahar Ben, et al.
Published: (2026)
by: Natan, Shahar Ben, et al.
Published: (2026)
Is Crowdsourcing Breaking Your Bank? Cost-Effective Fine-Tuning of Pre-trained Language Models with Proximal Policy Optimization
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
by: Kirchhof, Michael, et al.
Published: (2025)
by: Kirchhof, Michael, et al.
Published: (2025)
Self-Blinding and Counterfactual Self-Simulation Mitigate Biases and Sycophancy in Large Language Models
by: Christian, Brian, et al.
Published: (2026)
by: Christian, Brian, et al.
Published: (2026)
Understanding Knowledge Drift in LLMs through Misinformation
by: Fastowski, Alina, et al.
Published: (2024)
by: Fastowski, Alina, et al.
Published: (2024)
Emergent Abilities in Large Language Models: A Survey
by: Berti, Leonardo, et al.
Published: (2025)
by: Berti, Leonardo, et al.
Published: (2025)
"Check My Work?": Measuring Sycophancy in a Simulated Educational Context
by: Arvin, Chuck
Published: (2025)
by: Arvin, Chuck
Published: (2025)
SWAY: A Counterfactual Computational Linguistic Approach to Measuring and Mitigating Sycophancy
by: Bhalla, Joy, et al.
Published: (2026)
by: Bhalla, Joy, et al.
Published: (2026)
Not All Features Deserve Attention: Graph-Guided Dependency Learning for Tabular Data Generation with Language Models
by: Zhang, Zheyu, et al.
Published: (2025)
by: Zhang, Zheyu, et al.
Published: (2025)
HerO at AVeriTeC: The Herd of Open Large Language Models for Verifying Real-World Claims
by: Yoon, Yejun, et al.
Published: (2024)
by: Yoon, Yejun, et al.
Published: (2024)
Communication Bias in Large Language Models: A Regulatory Perspective
by: Kuenzler, Adrian, et al.
Published: (2025)
by: Kuenzler, Adrian, et al.
Published: (2025)
Adoption of Explainable Natural Language Processing: Perspectives from Industry and Academia on Practices and Challenges
by: Dhaini, Mahdi, et al.
Published: (2025)
by: Dhaini, Mahdi, et al.
Published: (2025)
From Confidence to Collapse in LLM Factual Robustness
by: Fastowski, Alina, et al.
Published: (2025)
by: Fastowski, Alina, et al.
Published: (2025)
Consolidating Rewarded Perturbations for LLM Post-Training
by: Zhang, Zheyu, et al.
Published: (2026)
by: Zhang, Zheyu, et al.
Published: (2026)
RAZOR: Sharpening Knowledge by Cutting Bias with Unsupervised Text Rewriting
by: Yang, Shuo, et al.
Published: (2024)
by: Yang, Shuo, et al.
Published: (2024)
Enhancing Fairness through Reweighting: A Path to Attain the Sufficiency Rule
by: Zhao, Xuan, et al.
Published: (2024)
by: Zhao, Xuan, et al.
Published: (2024)
CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models
by: Kocak, Aysenur, et al.
Published: (2025)
by: Kocak, Aysenur, et al.
Published: (2025)
Taking the Next Step with Generative Artificial Intelligence: The Transformative Role of Multimodal Large Language Models in Science Education
by: Bewersdorff, Arne, et al.
Published: (2024)
by: Bewersdorff, Arne, et al.
Published: (2024)
WinoQueer: A Community-in-the-Loop Benchmark for Anti-LGBTQ+ Bias in Large Language Models
by: Felkner, Virginia K., et al.
Published: (2023)
by: Felkner, Virginia K., et al.
Published: (2023)
Probabilistic Aggregation and Targeted Embedding Optimization for Collective Moral Reasoning in Large Language Models
by: Yuan, Chenchen, et al.
Published: (2025)
by: Yuan, Chenchen, et al.
Published: (2025)
When Explainability Meets Privacy: An Investigation at the Intersection of Post-hoc Explainability and Differential Privacy in the Context of Natural Language Processing
by: Dhaini, Mahdi, et al.
Published: (2025)
by: Dhaini, Mahdi, et al.
Published: (2025)
I Prefer not to Say: Protecting User Consent in Models with Optional Personal Data
by: Leemann, Tobias, et al.
Published: (2022)
by: Leemann, Tobias, et al.
Published: (2022)
ClaimVer: Explainable Claim-Level Verification and Evidence Attribution of Text Through Knowledge Graphs
by: Dammu, Preetam Prabhu Srikar, et al.
Published: (2024)
by: Dammu, Preetam Prabhu Srikar, et al.
Published: (2024)
Dutch Metaphor Extraction from Cancer Patients' Interviews and Forum Data using LLMs and Human in the Loop
by: Han, Lifeng, et al.
Published: (2025)
by: Han, Lifeng, et al.
Published: (2025)
Do Large Language Models Get Caught in Hofstadter-Mobius Loops?
by: Hryszko, Jaroslaw
Published: (2026)
by: Hryszko, Jaroslaw
Published: (2026)
Leveraging Large Language Models for Predictive Analysis of Human Misery
by: Seal, Bishanka, et al.
Published: (2025)
by: Seal, Bishanka, et al.
Published: (2025)
Europe's AI Imperative -- A Pragmatic Blueprint for Global Tech Leadership
by: Kasneci, Gjergji, et al.
Published: (2025)
by: Kasneci, Gjergji, et al.
Published: (2025)
Evaluating the Performance of Large Language Models in Scientific Claim Detection and Classification
by: Faruk, Tanjim Bin
Published: (2024)
by: Faruk, Tanjim Bin
Published: (2024)
Research Community Perspectives on "Intelligence" and Large Language Models
by: Højer, Bertram, et al.
Published: (2025)
by: Højer, Bertram, et al.
Published: (2025)
Strategic Insights in Human and Large Language Model Tactics at Word Guessing Games
by: Rikters, Matīss, et al.
Published: (2024)
by: Rikters, Matīss, et al.
Published: (2024)
SocialGaze: Improving the Integration of Human Social Norms in Large Language Models
by: Vijjini, Anvesh Rao, et al.
Published: (2024)
by: Vijjini, Anvesh Rao, et al.
Published: (2024)
Large Language Models' Accuracy in Emulating Human Experts' Evaluation of Public Sentiments about Heated Tobacco Products on Social Media
by: Kim, Kwanho, et al.
Published: (2025)
by: Kim, Kwanho, et al.
Published: (2025)
Attention Mechanisms Don't Learn Additive Models: Rethinking Feature Importance for Transformers
by: Leemann, Tobias, et al.
Published: (2024)
by: Leemann, Tobias, et al.
Published: (2024)
IDEAlign: Comparing Large Language Models to Human Experts in Open-ended Interpretive Annotations
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
Oversight Structures for Agentic AI in Public-Sector Organizations
by: Schmitz, Chris, et al.
Published: (2025)
by: Schmitz, Chris, et al.
Published: (2025)
Human-Level and Beyond: Benchmarking Large Language Models Against Clinical Pharmacists in Prescription Review
by: Yang, Yan, et al.
Published: (2025)
by: Yang, Yan, et al.
Published: (2025)
Similar Items
-
GermanPartiesQA: Benchmarking Commercial Large Language Models and AI Companions for Political Alignment and Sycophancy
by: Batzner, Jan, et al.
Published: (2024) -
Whose Personae? Synthetic Persona Experiments in LLM Research and Pathways to Transparency
by: Batzner, Jan, et al.
Published: (2025) -
The Language of Trauma: Modeling Traumatic Event Descriptions Across Domains with Explainable AI
by: Schirmer, Miriam, et al.
Published: (2024) -
Sycophancy is an Educational Safety Risk: Why LLM Tutors Need Sycophancy Benchmarks
by: Kasneci, Enkelejda, et al.
Published: (2026) -
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
by: Natan, Shahar Ben, et al.
Published: (2026)