Moral Sycophancy in Vision Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Rabby, Shadman, Papon, Md. Hefzul Hossain, Ahmed, Sabbir, Arif, Nokimul Hasan, Rahman, A. B. M. Ashikur, Ahmad, Irfan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Preemptive Hallucination Reduction: An Input-Level Approach for Multimodal Language Model
by: Arif, Nokimul Hasan, et al.
Published: (2025)
by: Arif, Nokimul Hasan, et al.
Published: (2025)
Experimental Comparison of Light-Weight and Deep CNN Models Across Diverse Datasets
by: Papon, Md. Hefzul Hossain, et al.
Published: (2026)
by: Papon, Md. Hefzul Hossain, et al.
Published: (2026)
PENDULUM: A Benchmark for Assessing Sycophancy in Multimodal Large Language Models
by: Rahman, A. B. M. Ashikur, et al.
Published: (2025)
by: Rahman, A. B. M. Ashikur, et al.
Published: (2025)
Conjunctive Prompt Attacks in Multi-Agent LLM Systems
by: Arif, Nokimul Hasan, et al.
Published: (2026)
by: Arif, Nokimul Hasan, et al.
Published: (2026)
Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language Models
by: Samin, Niamul Hassan, et al.
Published: (2026)
by: Samin, Niamul Hassan, et al.
Published: (2026)
Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in Long-Horizon Vision-Language Models
by: Rahman, Md Ashikur, et al.
Published: (2026)
by: Rahman, Md Ashikur, et al.
Published: (2026)
To Agree or To Be Right? The Grounding-Sycophancy Tradeoff in Medical Vision-Language Models
by: Aranya, OFM Riaz Rahman, et al.
Published: (2026)
by: Aranya, OFM Riaz Rahman, et al.
Published: (2026)
Securing Vision-Language Models with a Robust Encoder Against Jailbreak and Adversarial Attacks
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
Benchmarking and Mitigating Sycophancy in Medical Vision Language Models
by: Xu, Juangui, et al.
Published: (2025)
by: Xu, Juangui, et al.
Published: (2025)
Hybrid Deep Learning Approach for Coupled Demand Forecasting and Supply Chain Optimization
by: Nadia, Nusrat Yasmin, et al.
Published: (2026)
by: Nadia, Nusrat Yasmin, et al.
Published: (2026)
Pressure, What Pressure? Sycophancy Disentanglement in Language Models via Reward Decomposition
by: Mohsin, Muhammad Ahmed, et al.
Published: (2026)
by: Mohsin, Muhammad Ahmed, et al.
Published: (2026)
Restoring Rhythm: Punctuation Restoration Using Transformer Models for Bangla, A Low-Resource Language
by: Mamun, Md Obyedullahil, et al.
Published: (2025)
by: Mamun, Md Obyedullahil, et al.
Published: (2025)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026)
by: Erfan, Md, et al.
Published: (2026)
Sim-CLIP: Unsupervised Siamese Adversarial Fine-Tuning for Robust and Semantically-Rich Vision-Language Models
by: Hossain, Md Zarif, et al.
Published: (2024)
by: Hossain, Md Zarif, et al.
Published: (2024)
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
by: Natan, Shahar Ben, et al.
Published: (2026)
by: Natan, Shahar Ben, et al.
Published: (2026)
EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models
by: Yuan, Botai, et al.
Published: (2025)
by: Yuan, Botai, et al.
Published: (2025)
FUSED-Net: Detecting Traffic Signs with Limited Data
by: Rahman, Md. Atiqur, et al.
Published: (2024)
by: Rahman, Md. Atiqur, et al.
Published: (2024)
When Helpfulness Becomes Sycophancy: Sycophancy is a Boundary Failure Between Social Alignment and Epistemic Integrity in Large Language Models
by: Li, Jiechen, et al.
Published: (2026)
by: Li, Jiechen, et al.
Published: (2026)
Sycophancy in Vision-Language Models: A Systematic Analysis and an Inference-Time Mitigation Framework
by: Zhao, Yunpu, et al.
Published: (2024)
by: Zhao, Yunpu, et al.
Published: (2024)
Sycophancy in Large Language Models: Causes and Mitigations
by: Malmqvist, Lars
Published: (2024)
by: Malmqvist, Lars
Published: (2024)
NeuroSym-BioCAT: Leveraging Neuro-Symbolic Methods for Biomedical Scholarly Document Categorization and Question Answering
by: Zamil, Parvez, et al.
Published: (2024)
by: Zamil, Parvez, et al.
Published: (2024)
DL$^3$M: A Vision-to-Language Framework for Expert-Level Medical Reasoning through Deep Learning and Large Language Models
by: Hasan, Md. Najib, et al.
Published: (2025)
by: Hasan, Md. Najib, et al.
Published: (2025)
Missile detection and destruction robot using detection algorithm
by: Siam, Md Kamrul, et al.
Published: (2024)
by: Siam, Md Kamrul, et al.
Published: (2024)
Overlapping Community Detection using Dynamic Dilated Aggregation in Deep Residual GCN
by: Muttakin, Md Nurul, et al.
Published: (2022)
by: Muttakin, Md Nurul, et al.
Published: (2022)
Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models
by: Denison, Carson, et al.
Published: (2024)
by: Denison, Carson, et al.
Published: (2024)
Accounting for Sycophancy in Language Model Uncertainty Estimation
by: Sicilia, Anthony, et al.
Published: (2024)
by: Sicilia, Anthony, et al.
Published: (2024)
Towards Understanding Sycophancy in Language Models
by: Sharma, Mrinank, et al.
Published: (2023)
by: Sharma, Mrinank, et al.
Published: (2023)
Beyond Symbolic Solving: Multi Chain-of-Thought Voting for Geometric Reasoning in Large Language Models
by: Siddique, Md. Abu Bakor, et al.
Published: (2026)
by: Siddique, Md. Abu Bakor, et al.
Published: (2026)
A Bidirectional Siamese Recurrent Neural Network for Accurate Gait Recognition Using Body Landmarks
by: Progga, Proma Hossain, et al.
Published: (2024)
by: Progga, Proma Hossain, et al.
Published: (2024)
CLIN-LLM: A Safety-Constrained Hybrid Framework for Clinical Diagnosis and Treatment Generation
by: Hasan, Md. Mehedi, et al.
Published: (2025)
by: Hasan, Md. Mehedi, et al.
Published: (2025)
Sentra-Guard: A Real-Time Multilingual Defense Against Adversarial LLM Prompts
by: Hasan, Md. Mehedi, et al.
Published: (2025)
by: Hasan, Md. Mehedi, et al.
Published: (2025)
Fingerprinting Deep Learning Models via Network Traffic Patterns in Federated Learning
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)
by: Shuvo, Md Nahid Hasan, et al.
Published: (2025)
A Large Language Model-Supported Threat Modeling Framework for Transportation Cyber-Physical Systems
by: Salek, M Sabbir, et al.
Published: (2025)
by: Salek, M Sabbir, et al.
Published: (2025)
A Cascaded Architecture for Extractive Summarization of Multimedia Content via Audio-to-Text Alignment
by: Hossain, Tanzir, et al.
Published: (2025)
by: Hossain, Tanzir, et al.
Published: (2025)
How RLHF Amplifies Sycophancy
by: Shapira, Itai, et al.
Published: (2026)
by: Shapira, Itai, et al.
Published: (2026)
SycEval: Evaluating LLM Sycophancy
by: Fanous, Aaron, et al.
Published: (2025)
by: Fanous, Aaron, et al.
Published: (2025)
Enhancement of Bengali OCR by Specialized Models and Advanced Techniques for Diverse Document Types
by: Rabby, AKM Shahariar Azad, et al.
Published: (2024)
by: Rabby, AKM Shahariar Azad, et al.
Published: (2024)
A Survey on Causal Discovery Methods for I.I.D. and Time Series Data
by: Hasan, Uzma, et al.
Published: (2023)
by: Hasan, Uzma, et al.
Published: (2023)
Investigating the Influence of Language on Sycophantic Behavior of Multilingual LLMs
by: Aldahlawi, Bayan Abdullah, et al.
Published: (2026)
by: Aldahlawi, Bayan Abdullah, et al.
Published: (2026)
An Efficient Deep Learning Framework for Brain Stroke Diagnosis Using Computed Tomography Images
by: Hossen, Md. Sabbir, et al.
Published: (2025)
by: Hossen, Md. Sabbir, et al.
Published: (2025)
Similar Items
-
Preemptive Hallucination Reduction: An Input-Level Approach for Multimodal Language Model
by: Arif, Nokimul Hasan, et al.
Published: (2025) -
Experimental Comparison of Light-Weight and Deep CNN Models Across Diverse Datasets
by: Papon, Md. Hefzul Hossain, et al.
Published: (2026) -
PENDULUM: A Benchmark for Assessing Sycophancy in Multimodal Large Language Models
by: Rahman, A. B. M. Ashikur, et al.
Published: (2025) -
Conjunctive Prompt Attacks in Multi-Agent LLM Systems
by: Arif, Nokimul Hasan, et al.
Published: (2026) -
Beyond Dominant Patches: Spatial Credit Redistribution For Grounded Vision-Language Models
by: Samin, Niamul Hassan, et al.
Published: (2026)