Benchmarking and Mitigating Sycophancy in Medical Vision Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Juangui, Guo, Zikun, Lv, Jingwei, Lin, Hongbin, Yang, Shu, Wen, Jun, Wang, Di, Hu, Lijie |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
di: Zhou, Wenrui, et al.
Pubblicazione: (2025)
di: Zhou, Wenrui, et al.
Pubblicazione: (2025)
EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models
di: Yuan, Botai, et al.
Pubblicazione: (2025)
di: Yuan, Botai, et al.
Pubblicazione: (2025)
To Agree or To Be Right? The Grounding-Sycophancy Tradeoff in Medical Vision-Language Models
di: Aranya, OFM Riaz Rahman, et al.
Pubblicazione: (2026)
di: Aranya, OFM Riaz Rahman, et al.
Pubblicazione: (2026)
Stable Vision Concept Transformers for Medical Diagnosis
di: Hu, Lijie, et al.
Pubblicazione: (2025)
di: Hu, Lijie, et al.
Pubblicazione: (2025)
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images
di: You, Liangliang, et al.
Pubblicazione: (2025)
di: You, Liangliang, et al.
Pubblicazione: (2025)
Graph-Guided Dual-Level Augmentation for 3D Scene Segmentation
di: Lin, Hongbin, et al.
Pubblicazione: (2025)
di: Lin, Hongbin, et al.
Pubblicazione: (2025)
Tracing and Mitigating Hallucinations in Multimodal LLMs via Dynamic Attention Localization
di: Yang, Tiancheng, et al.
Pubblicazione: (2025)
di: Yang, Tiancheng, et al.
Pubblicazione: (2025)
MedFM-Robust: Benchmarking Robustness of Medical Foundation Models
di: Cui, Xiangxiang, et al.
Pubblicazione: (2026)
di: Cui, Xiangxiang, et al.
Pubblicazione: (2026)
HalluCXR: Benchmarking and Mitigating Hallucinations in Medical Vision-Language Models for Chest Radiograph Interpretation
di: Wang, Haoyu, et al.
Pubblicazione: (2026)
di: Wang, Haoyu, et al.
Pubblicazione: (2026)
ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
di: Chen, Junzhe, et al.
Pubblicazione: (2024)
MedHallTune: An Instruction-Tuning Benchmark for Mitigating Medical Hallucination in Vision-Language Models
di: Yan, Qiao, et al.
Pubblicazione: (2025)
di: Yan, Qiao, et al.
Pubblicazione: (2025)
Identifying and Mitigating Position Bias of Multi-image Vision-Language Models
di: Tian, Xinyu, et al.
Pubblicazione: (2025)
di: Tian, Xinyu, et al.
Pubblicazione: (2025)
CapNav: Benchmarking Vision Language Models on Capability-conditioned Indoor Navigation
di: Su, Xia, et al.
Pubblicazione: (2026)
di: Su, Xia, et al.
Pubblicazione: (2026)
MedHEval: Benchmarking Hallucinations and Mitigation Strategies in Medical Large Vision-Language Models
di: Chang, Aofei, et al.
Pubblicazione: (2025)
di: Chang, Aofei, et al.
Pubblicazione: (2025)
SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals
di: Lin, Zihang, et al.
Pubblicazione: (2026)
di: Lin, Zihang, et al.
Pubblicazione: (2026)
Robust Fairness Vision-Language Learning for Medical Image Analysis
di: Bansal, Sparsh, et al.
Pubblicazione: (2025)
di: Bansal, Sparsh, et al.
Pubblicazione: (2025)
Visual Self-Fulfilling Alignment: Shaping Safety-Oriented Personas via Threat-Related Images
di: Yang, Qishun, et al.
Pubblicazione: (2026)
di: Yang, Qishun, et al.
Pubblicazione: (2026)
Benchmarking and Mitigating MCQA Selection Bias of Large Vision-Language Models
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
di: Atabuzzaman, Md., et al.
Pubblicazione: (2025)
Towards Multi-dimensional Explanation Alignment for Medical Classification
di: Hu, Lijie, et al.
Pubblicazione: (2024)
di: Hu, Lijie, et al.
Pubblicazione: (2024)
Cross-Modal Attention Analysis and Optimization in Vision-Language Models: A Study on Visual Reliability
di: Zhou, Lijie
Pubblicazione: (2026)
di: Zhou, Lijie
Pubblicazione: (2026)
LLMC+: Benchmarking Vision-Language Model Compression with a Plug-and-play Toolkit
di: Lv, Chengtao, et al.
Pubblicazione: (2025)
di: Lv, Chengtao, et al.
Pubblicazione: (2025)
Light as Deception: GPT-driven Natural Relighting Against Vision-Language Pre-training Models
di: Yang, Ying, et al.
Pubblicazione: (2025)
di: Yang, Ying, et al.
Pubblicazione: (2025)
Spatiotemporal Sycophancy: Negation-Based Gaslighting in Video Large Language Models
di: Tang, Ziyao, et al.
Pubblicazione: (2026)
di: Tang, Ziyao, et al.
Pubblicazione: (2026)
VLRMBench: A Comprehensive and Challenging Benchmark for Vision-Language Reward Models
di: Ruan, Jiacheng, et al.
Pubblicazione: (2025)
di: Ruan, Jiacheng, et al.
Pubblicazione: (2025)
When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models
di: Yakun, Cui, et al.
Pubblicazione: (2026)
di: Yakun, Cui, et al.
Pubblicazione: (2026)
MambaVLT: Time-Evolving Multimodal State Space Model for Vision-Language Tracking
di: Liu, Xinqi, et al.
Pubblicazione: (2024)
di: Liu, Xinqi, et al.
Pubblicazione: (2024)
debiaSAE: Benchmarking and Mitigating Vision-Language Model Bias
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
Prompting Medical Vision-Language Models to Mitigate Diagnosis Bias by Generating Realistic Dermoscopic Images
di: Munia, Nusrat, et al.
Pubblicazione: (2025)
di: Munia, Nusrat, et al.
Pubblicazione: (2025)
BenchX: A Unified Benchmark Framework for Medical Vision-Language Pretraining on Chest X-Rays
di: Zhou, Yang, et al.
Pubblicazione: (2024)
di: Zhou, Yang, et al.
Pubblicazione: (2024)
Gaze-directed Vision GNN for Mitigating Shortcut Learning in Medical Image
di: Wu, Shaoxuan, et al.
Pubblicazione: (2024)
di: Wu, Shaoxuan, et al.
Pubblicazione: (2024)
Uncertainty-Driven Expert Control: Enhancing the Reliability of Medical Vision-Language Models
di: Liang, Xiao, et al.
Pubblicazione: (2025)
di: Liang, Xiao, et al.
Pubblicazione: (2025)
Hulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language Understanding
di: Jiang, Songtao, et al.
Pubblicazione: (2025)
di: Jiang, Songtao, et al.
Pubblicazione: (2025)
Pointing to a Llama and Call it a Camel: On the Sycophancy of Multimodal Large Language Models
di: Pi, Renjie, et al.
Pubblicazione: (2025)
di: Pi, Renjie, et al.
Pubblicazione: (2025)
YARD: Y-Architecture Register Decoding for Efficient Hallucination Mitigation in Large Vision-Language Models
di: Chen, Ting, et al.
Pubblicazione: (2026)
di: Chen, Ting, et al.
Pubblicazione: (2026)
Backdooring Vision-Language Models with Out-Of-Distribution Data
di: Lyu, Weimin, et al.
Pubblicazione: (2024)
di: Lyu, Weimin, et al.
Pubblicazione: (2024)
Med-VCD: Mitigating Hallucination for Medical Large Vision Language Models through Visual Contrastive Decoding
di: Mahdavi, Zahra, et al.
Pubblicazione: (2025)
di: Mahdavi, Zahra, et al.
Pubblicazione: (2025)
Harnessing Vision-Language Pretrained Models with Temporal-Aware Adaptation for Referring Video Object Segmentation
di: Zhou, Zikun, et al.
Pubblicazione: (2024)
di: Zhou, Zikun, et al.
Pubblicazione: (2024)
Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration
di: Xu, Yuanzhi, et al.
Pubblicazione: (2026)
di: Xu, Yuanzhi, et al.
Pubblicazione: (2026)
OmniEarth: A Benchmark for Evaluating Vision-Language Models in Geospatial Tasks
di: Fu, Ronghao, et al.
Pubblicazione: (2026)
di: Fu, Ronghao, et al.
Pubblicazione: (2026)
Meta-Adapter: An Online Few-shot Learner for Vision-Language Model
di: Cheng, Cheng, et al.
Pubblicazione: (2023)
di: Cheng, Cheng, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
di: Zhou, Wenrui, et al.
Pubblicazione: (2025) -
EchoBench: Benchmarking Sycophancy in Medical Large Vision-Language Models
di: Yuan, Botai, et al.
Pubblicazione: (2025) -
To Agree or To Be Right? The Grounding-Sycophancy Tradeoff in Medical Vision-Language Models
di: Aranya, OFM Riaz Rahman, et al.
Pubblicazione: (2026) -
Stable Vision Concept Transformers for Medical Diagnosis
di: Hu, Lijie, et al.
Pubblicazione: (2025) -
Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images
di: You, Liangliang, et al.
Pubblicazione: (2025)