The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Zhenyu, Balagopalan, Aparna, Agrawal, Adi, Yergasheva, Dilshoda, Alshikh, Waseem, Bikel, Daniel M. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Accurate Failure Prediction in Agents Does Not Imply Effective Failure Prevention
by: Vasudev, Rakshith, et al.
Published: (2026)
by: Vasudev, Rakshith, et al.
Published: (2026)
Shorthand for Thought: Compressing LLM Reasoning via Entropy-Guided Supertokens
by: Zhao, Zhenyu, et al.
Published: (2026)
by: Zhao, Zhenyu, et al.
Published: (2026)
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty
by: Guo, Kevin H., et al.
Published: (2026)
by: Guo, Kevin H., et al.
Published: (2026)
Post-training an LLM for RAG? Train on Self-Generated Demonstrations
by: Finlayson, Matthew, et al.
Published: (2025)
by: Finlayson, Matthew, et al.
Published: (2025)
Consistency Training Helps Stop Sycophancy and Jailbreaks
by: Irpan, Alex, et al.
Published: (2025)
by: Irpan, Alex, et al.
Published: (2025)
PARROT: Persuasion and Agreement Robustness Rating of Output Truth -- A Sycophancy Robustness Benchmark for LLMs
by: Çelebi, Yusuf, et al.
Published: (2025)
by: Çelebi, Yusuf, et al.
Published: (2025)
Fixed Budget is No Harder Than Fixed Confidence in Best-Arm Identification up to Logarithmic Factors
by: Balagopalan, Kapilan, et al.
Published: (2026)
by: Balagopalan, Kapilan, et al.
Published: (2026)
Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy
by: Kumarappan, Adarsh, et al.
Published: (2026)
by: Kumarappan, Adarsh, et al.
Published: (2026)
Behavioural Effects of Agentic Messaging: A Case Study on a Financial Service Application
by: Jeunen, Olivier, et al.
Published: (2025)
by: Jeunen, Olivier, et al.
Published: (2025)
AMA-Bench: Evaluating Long-Horizon Memory for Agentic Applications
by: Zhao, Yujie, et al.
Published: (2026)
by: Zhao, Yujie, et al.
Published: (2026)
Phase-Aware Deep Learning with Complex-Valued CNNs for Audio Signal Applications
by: Agrawal, Naman
Published: (2025)
by: Agrawal, Naman
Published: (2025)
Backtracking Improves Generation Safety
by: Zhang, Yiming, et al.
Published: (2024)
by: Zhang, Yiming, et al.
Published: (2024)
BrokenMath: A Benchmark for Sycophancy in Theorem Proving with LLMs
by: Petrov, Ivo, et al.
Published: (2025)
by: Petrov, Ivo, et al.
Published: (2025)
LLM4Cov: Execution-Aware Agentic Learning for High-coverage Testbench Generation
by: Zhang, Hejia, et al.
Published: (2026)
by: Zhang, Hejia, et al.
Published: (2026)
Model Agreement via Anchoring
by: Eaton, Eric, et al.
Published: (2026)
by: Eaton, Eric, et al.
Published: (2026)
Significativity Indices for Agreement Values
by: Casagrande, Alberto, et al.
Published: (2025)
by: Casagrande, Alberto, et al.
Published: (2025)
Among Us: A Sandbox for Measuring and Detecting Agentic Deception
by: Golechha, Satvik, et al.
Published: (2025)
by: Golechha, Satvik, et al.
Published: (2025)
Beyond Next Word Prediction: Developing Comprehensive Evaluation Frameworks for measuring LLM performance on real world applications
by: Agrawal, Vishakha, et al.
Published: (2025)
by: Agrawal, Vishakha, et al.
Published: (2025)
Towards Understanding Sycophancy in Language Models
by: Sharma, Mrinank, et al.
Published: (2023)
by: Sharma, Mrinank, et al.
Published: (2023)
Agentic Unlearning: When LLM Agent Meets Machine Unlearning
by: Wang, Bin, et al.
Published: (2026)
by: Wang, Bin, et al.
Published: (2026)
TAGAL: Tabular Data Generation using Agentic LLM Methods
by: Ronval, Benoît, et al.
Published: (2025)
by: Ronval, Benoît, et al.
Published: (2025)
Multiple-Resolution Tokenization for Time Series Forecasting with an Application to Pricing
by: Peršak, Egon, et al.
Published: (2024)
by: Peršak, Egon, et al.
Published: (2024)
Addressing LLM Diversity by Infusing Random Concepts
by: Agrawal, Pulin, et al.
Published: (2026)
by: Agrawal, Pulin, et al.
Published: (2026)
Predicting Stock Price Movement with LLM-Enhanced Tweet Emotion Analysis
by: Vuong, An, et al.
Published: (2025)
by: Vuong, An, et al.
Published: (2025)
IRL for Restless Multi-Armed Bandits with Applications in Maternal and Child Health
by: Jain, Gauri, et al.
Published: (2024)
by: Jain, Gauri, et al.
Published: (2024)
Multi-View Encoders for Performance Prediction in LLM-Based Agentic Workflows
by: Trirat, Patara, et al.
Published: (2025)
by: Trirat, Patara, et al.
Published: (2025)
Forecasting Commodity Price Shocks Using Temporal and Semantic Fusion of Prices Signals and Agentic Generative AI Extracted Economic News
by: Ghali, Mohammed-Khalil, et al.
Published: (2025)
by: Ghali, Mohammed-Khalil, et al.
Published: (2025)
Agentic AI: A Comprehensive Survey of Architectures, Applications, and Future Directions
by: Ali, Mohamad Abou, et al.
Published: (2025)
by: Ali, Mohamad Abou, et al.
Published: (2025)
Adaptive Few-Shot Learning (AFSL): Tackling Data Scarcity with Stability, Robustness, and Versatility
by: Agrawal, Rishabh
Published: (2025)
by: Agrawal, Rishabh
Published: (2025)
A Generalized Bisimulation Metric of State Similarity between Markov Decision Processes: From Theoretical Propositions to Applications
by: Tao, Zhenyu, et al.
Published: (2025)
by: Tao, Zhenyu, et al.
Published: (2025)
Agentic Skill Discovery
by: Zhao, Xufeng, et al.
Published: (2024)
by: Zhao, Xufeng, et al.
Published: (2024)
Agentic Uncertainty Reveals Agentic Overconfidence
by: Kaddour, Jean, et al.
Published: (2026)
by: Kaddour, Jean, et al.
Published: (2026)
AgenticPay: A Multi-Agent LLM Negotiation System for Buyer-Seller Transactions
by: Liu, Xianyang, et al.
Published: (2026)
by: Liu, Xianyang, et al.
Published: (2026)
Calibration Collapse Under Sycophancy Fine-Tuning: How Reward Hacking Breaks Uncertainty Quantification in LLMs
by: Sahoo, Subramanyam
Published: (2026)
by: Sahoo, Subramanyam
Published: (2026)
Skill Reuse as Compression in Agentic RL
by: Xu, Zhikun, et al.
Published: (2026)
by: Xu, Zhikun, et al.
Published: (2026)
InterCorpRel-LLM: Enhancing Financial Relational Understanding with Graph-Language Models
by: Sun, Qianyou, et al.
Published: (2025)
by: Sun, Qianyou, et al.
Published: (2025)
Test-Time Adaptation Induces Stronger Accuracy and Agreement-on-the-Line
by: Kim, Eungyeup, et al.
Published: (2023)
by: Kim, Eungyeup, et al.
Published: (2023)
Sycophantic Anchors: Localizing and Quantifying User Agreement in Reasoning Models
by: Duszenko, Jacek
Published: (2026)
by: Duszenko, Jacek
Published: (2026)
On Measuring Unnoticeability of Graph Adversarial Attacks: Observations, New Measure, and Applications
by: Jo, Hyeonsoo, et al.
Published: (2025)
by: Jo, Hyeonsoo, et al.
Published: (2025)
Have the VLMs Lost Confidence? A Study of Sycophancy in VLMs
by: Li, Shuo, et al.
Published: (2024)
by: Li, Shuo, et al.
Published: (2024)
Similar Items
-
Accurate Failure Prediction in Agents Does Not Imply Effective Failure Prevention
by: Vasudev, Rakshith, et al.
Published: (2026) -
Shorthand for Thought: Compressing LLM Reasoning via Entropy-Guided Supertokens
by: Zhao, Zhenyu, et al.
Published: (2026) -
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty
by: Guo, Kevin H., et al.
Published: (2026) -
Post-training an LLM for RAG? Train on Self-Generated Demonstrations
by: Finlayson, Matthew, et al.
Published: (2025) -
Consistency Training Helps Stop Sycophancy and Jailbreaks
by: Irpan, Alex, et al.
Published: (2025)