Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Byun, Ji Young, Park, Young-Jin, Corbeil, Jean-Philippe, Abacha, Asma Ben |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Improving Medical VQA through Trajectory-Aware Process Supervision
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026)
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026)
SURE-VQA: Systematic Understanding of Robustness Evaluation in Medical VQA Tasks
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024)
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024)
Known Meets Unknown: Mitigating Overconfidence in Open Set Recognition
von: Zhao, Dongdong, et al.
Veröffentlicht: (2025)
von: Zhao, Dongdong, et al.
Veröffentlicht: (2025)
FilterRAG: Zero-Shot Informed Retrieval-Augmented Generation to Mitigate Hallucinations in VQA
von: Sarwar, Nobin
Veröffentlicht: (2025)
von: Sarwar, Nobin
Veröffentlicht: (2025)
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
von: Kolli, Govinda, et al.
Veröffentlicht: (2026)
von: Kolli, Govinda, et al.
Veröffentlicht: (2026)
Look Closer! An Adversarial Parametric Editing Framework for Hallucination Mitigation in VLMs
von: Hu, Jiayu, et al.
Veröffentlicht: (2025)
von: Hu, Jiayu, et al.
Veröffentlicht: (2025)
Efficient High-Resolution Image Editing with Hallucination-Aware Loss and Adaptive Tiling
von: Kwon, Young D., et al.
Veröffentlicht: (2025)
von: Kwon, Young D., et al.
Veröffentlicht: (2025)
ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models
von: Park, Yeji, et al.
Veröffentlicht: (2024)
von: Park, Yeji, et al.
Veröffentlicht: (2024)
Optimal Transport-Induced Samples against Out-of-Distribution Overconfidence
von: Tang, Keke, et al.
Veröffentlicht: (2026)
von: Tang, Keke, et al.
Veröffentlicht: (2026)
Mitigating Diffusion Model Hallucinations with Dynamic Guidance
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
Bridging the Semantic Gaps: Improving Medical VQA Consistency with LLM-Augmented Question Sets
von: Ma, Yongpei, et al.
Veröffentlicht: (2025)
von: Ma, Yongpei, et al.
Veröffentlicht: (2025)
Refine and Align: Confidence Calibration through Multi-Agent Interaction in VQA
von: Pandey, Ayush, et al.
Veröffentlicht: (2025)
von: Pandey, Ayush, et al.
Veröffentlicht: (2025)
Online Self-Calibration Against Hallucination in Vision-Language Models
von: Chen, Minghui, et al.
Veröffentlicht: (2026)
von: Chen, Minghui, et al.
Veröffentlicht: (2026)
Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations
von: Jiang, Nick, et al.
Veröffentlicht: (2024)
von: Jiang, Nick, et al.
Veröffentlicht: (2024)
Mitigating Algorithmic Bias in Multiclass CNN Classifications Using Causal Modeling
von: Byun, Min Sik, et al.
Veröffentlicht: (2025)
von: Byun, Min Sik, et al.
Veröffentlicht: (2025)
Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration
von: Xu, Yuanzhi, et al.
Veröffentlicht: (2026)
von: Xu, Yuanzhi, et al.
Veröffentlicht: (2026)
Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding
von: Tang, Feilong, et al.
Veröffentlicht: (2025)
von: Tang, Feilong, et al.
Veröffentlicht: (2025)
Locate-then-Sparsify: Attribution Guided Sparse Strategy for Visual Hallucination Mitigation
von: Dang, Tiantian, et al.
Veröffentlicht: (2026)
von: Dang, Tiantian, et al.
Veröffentlicht: (2026)
WildFireVQA: A Large-Scale Radiometric Thermal VQA Benchmark for Aerial Wildfire Monitoring
von: Habibpour, Mobin, et al.
Veröffentlicht: (2026)
von: Habibpour, Mobin, et al.
Veröffentlicht: (2026)
An Empirical Study Into What Matters for Calibrating Vision-Language Models
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
von: Tu, Weijie, et al.
Veröffentlicht: (2024)
Test-Time-Scaling for Zero-Shot Diagnosis with Visual-Language Reasoning
von: Byun, Ji Young, et al.
Veröffentlicht: (2025)
von: Byun, Ji Young, et al.
Veröffentlicht: (2025)
MaskMedPaint: Masked Medical Image Inpainting with Diffusion Models for Mitigation of Spurious Correlations
von: Jin, Qixuan, et al.
Veröffentlicht: (2024)
von: Jin, Qixuan, et al.
Veröffentlicht: (2024)
Taking Shortcuts for Categorical VQA Using Super Neurons
von: Musacchio, Pierre, et al.
Veröffentlicht: (2026)
von: Musacchio, Pierre, et al.
Veröffentlicht: (2026)
Mitigating Hallucinations in Vision-Language Models through Image-Guided Head Suppression
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2025)
von: Sarkar, Sreetama, et al.
Veröffentlicht: (2025)
VideoHallu: Evaluating and Mitigating Multi-modal Hallucinations on Synthetic Video Understanding
von: Li, Zongxia, et al.
Veröffentlicht: (2025)
von: Li, Zongxia, et al.
Veröffentlicht: (2025)
HaloProbe: Bayesian Detection and Mitigation of Object Hallucinations in Vision-Language Models
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
von: Zohrabi, Reihaneh, et al.
Veröffentlicht: (2026)
Interpolation of GEDI Biomass Estimates with Calibrated Uncertainty Quantification
von: Young, Robin, et al.
Veröffentlicht: (2026)
von: Young, Robin, et al.
Veröffentlicht: (2026)
Leveraging Programmatically Generated Synthetic Data for Differentially Private Diffusion Training
von: Choi, Yujin, et al.
Veröffentlicht: (2024)
von: Choi, Yujin, et al.
Veröffentlicht: (2024)
Disentanglement-Based Equivariant Learning for Compositional VQA
von: Du, Zhou, et al.
Veröffentlicht: (2026)
von: Du, Zhou, et al.
Veröffentlicht: (2026)
Unexplored flaws in multiple-choice VQA evaluations
von: Rosenthal, Fabio, et al.
Veröffentlicht: (2025)
von: Rosenthal, Fabio, et al.
Veröffentlicht: (2025)
BERT-VQA: Visual Question Answering on Plots
von: Vu, Tai, et al.
Veröffentlicht: (2025)
von: Vu, Tai, et al.
Veröffentlicht: (2025)
VQA-Levels: A Hierarchical Approach for Classifying Questions in VQA
von: Madaka, Madhuri Latha, et al.
Veröffentlicht: (2025)
von: Madaka, Madhuri Latha, et al.
Veröffentlicht: (2025)
Optimizing Calibration by Gaining Aware of Prediction Correctness
von: Liu, Yuchi, et al.
Veröffentlicht: (2024)
von: Liu, Yuchi, et al.
Veröffentlicht: (2024)
Reducing catastrophic forgetting of incremental learning in the absence of rehearsal memory with task-specific token
von: Choi, Young Jo, et al.
Veröffentlicht: (2024)
von: Choi, Young Jo, et al.
Veröffentlicht: (2024)
Laplacian Score Sharpening for Mitigating Hallucination in Diffusion Models
von: C, Barath Chandran., et al.
Veröffentlicht: (2025)
von: C, Barath Chandran., et al.
Veröffentlicht: (2025)
Hallucinatory Image Tokens: A Training-free EAZY Approach on Detecting and Mitigating Object Hallucinations in LVLMs
von: Che, Liwei, et al.
Veröffentlicht: (2025)
von: Che, Liwei, et al.
Veröffentlicht: (2025)
Multi-Task Learning for Visually Grounded Reasoning in Gastrointestinal VQA
von: Safwan, Itbaan, et al.
Veröffentlicht: (2025)
von: Safwan, Itbaan, et al.
Veröffentlicht: (2025)
Frequency-Calibrated Membership Inference Attacks on Medical Image Diffusion Models
von: Zhao, Xinkai, et al.
Veröffentlicht: (2025)
von: Zhao, Xinkai, et al.
Veröffentlicht: (2025)
Overconfidence is Key: Verbalized Uncertainty Evaluation in Large Language and Vision-Language Models
von: Groot, Tobias, et al.
Veröffentlicht: (2024)
von: Groot, Tobias, et al.
Veröffentlicht: (2024)
RadFlag: A Black-Box Hallucination Detection Method for Medical Vision Language Models
von: Zhang, Serena, et al.
Veröffentlicht: (2024)
von: Zhang, Serena, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Improving Medical VQA through Trajectory-Aware Process Supervision
von: Gulluk, Halil Ibrahim, et al.
Veröffentlicht: (2026) -
SURE-VQA: Systematic Understanding of Robustness Evaluation in Medical VQA Tasks
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024) -
Known Meets Unknown: Mitigating Overconfidence in Open Set Recognition
von: Zhao, Dongdong, et al.
Veröffentlicht: (2025) -
FilterRAG: Zero-Shot Informed Retrieval-Augmented Generation to Mitigate Hallucinations in VQA
von: Sarwar, Nobin
Veröffentlicht: (2025) -
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
von: Kolli, Govinda, et al.
Veröffentlicht: (2026)