Auditing Frontier Vision-Language Models for Trustworthy Medical VQA: Grounding Failures, Format Collapse, and Domain Adaptation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Xupeng, Shi, Binbin, Le, Chenqian, Yin, Qifu, Lin, Lang, Ni, Haowei, Gong, Ran, Li, Panfeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering
von: Chen, Xupeng, et al.
Veröffentlicht: (2026)
von: Chen, Xupeng, et al.
Veröffentlicht: (2026)
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
Enhancing Exchange Rate Forecasting with Explainable Deep Learning Models
von: Meng, Shuchen, et al.
Veröffentlicht: (2024)
von: Meng, Shuchen, et al.
Veröffentlicht: (2024)
Time Series Modeling for Heart Rate Prediction: From ARIMA to Transformers
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
von: Ni, Haowei, et al.
Veröffentlicht: (2024)
Unified and Semantically Grounded Domain Adaptation for Medical Image Segmentation
von: Wang, Xin, et al.
Veröffentlicht: (2025)
von: Wang, Xin, et al.
Veröffentlicht: (2025)
Evaluating Modern Approaches in 3D Scene Reconstruction: NeRF vs Gaussian-Based Methods
von: Zhou, Yiming, et al.
Veröffentlicht: (2024)
von: Zhou, Yiming, et al.
Veröffentlicht: (2024)
VoxelFormer: Parameter-Efficient Multi-Subject Visual Decoding from fMRI
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
von: Le, Chenqian, et al.
Veröffentlicht: (2025)
On the Role of Visual Grounding in VQA
von: Reich, Daniel, et al.
Veröffentlicht: (2024)
von: Reich, Daniel, et al.
Veröffentlicht: (2024)
Explainable Information Retrieval in the Audit Domain
von: Frummet, Alexander, et al.
Veröffentlicht: (2025)
von: Frummet, Alexander, et al.
Veröffentlicht: (2025)
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence
von: Ma, Dongsheng, et al.
Veröffentlicht: (2026)
von: Ma, Dongsheng, et al.
Veröffentlicht: (2026)
Towards Trustworthy Unsupervised Domain Adaptation: A Representation Learning Perspective for Enhancing Robustness, Discrimination, and Generalization
von: Yin, Jia-Li, et al.
Veröffentlicht: (2024)
von: Yin, Jia-Li, et al.
Veröffentlicht: (2024)
MedXplain-VQA: Multi-Component Explainable Medical Visual Question Answering
von: Nguyen, Hai-Dang, et al.
Veröffentlicht: (2025)
von: Nguyen, Hai-Dang, et al.
Veröffentlicht: (2025)
SURE-VQA: Systematic Understanding of Robustness Evaluation in Medical VQA Tasks
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024)
von: Kahl, Kim-Celine, et al.
Veröffentlicht: (2024)
CARES: A Comprehensive Benchmark of Trustworthiness in Medical Vision Language Models
von: Xia, Peng, et al.
Veröffentlicht: (2024)
von: Xia, Peng, et al.
Veröffentlicht: (2024)
R-LLaVA: Improving Med-VQA Understanding through Visual Region of Interest
von: Chen, Xupeng, et al.
Veröffentlicht: (2024)
von: Chen, Xupeng, et al.
Veröffentlicht: (2024)
Tri-VQA: Triangular Reasoning Medical Visual Question Answering for Multi-Attribute Analysis
von: Fan, Lin, et al.
Veröffentlicht: (2024)
von: Fan, Lin, et al.
Veröffentlicht: (2024)
Tackling Dimensional Collapse toward Comprehensive Universal Domain Adaptation
von: Fang, Hung-Chieh, et al.
Veröffentlicht: (2024)
von: Fang, Hung-Chieh, et al.
Veröffentlicht: (2024)
Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models
von: Zhu, Wenhui, et al.
Veröffentlicht: (2025)
von: Zhu, Wenhui, et al.
Veröffentlicht: (2025)
Measuring Faithful and Plausible Visual Grounding in VQA
von: Reich, Daniel, et al.
Veröffentlicht: (2023)
von: Reich, Daniel, et al.
Veröffentlicht: (2023)
MedSkillAudit: A Domain-Specific Audit Framework for Medical Research Agent Skills
von: Hou, Yingyong, et al.
Veröffentlicht: (2026)
von: Hou, Yingyong, et al.
Veröffentlicht: (2026)
DeferredSeg: A Multi-Expert Deferral Framework for Trustworthy Medical Image Segmentation
von: Tian, Qiuyu, et al.
Veröffentlicht: (2026)
von: Tian, Qiuyu, et al.
Veröffentlicht: (2026)
Comparison of sEMG Encoding Accuracy Across Speech Modes Using Articulatory and Phoneme Features
von: Le, Chenqian, et al.
Veröffentlicht: (2026)
von: Le, Chenqian, et al.
Veröffentlicht: (2026)
Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables
von: Shen, Chen
Veröffentlicht: (2026)
von: Shen, Chen
Veröffentlicht: (2026)
RCS-SF-001 — Systemic Failure Under Recoverability Constraints: Cross-Domain Collapse and Unified Failure Pattern
von: Sanchez, Noelia, et al.
Veröffentlicht: (2026)
von: Sanchez, Noelia, et al.
Veröffentlicht: (2026)
Self-Supervised Anatomical Consistency Learning for Vision-Grounded Medical Report Generation
von: Yang, Longzhen, et al.
Veröffentlicht: (2025)
von: Yang, Longzhen, et al.
Veröffentlicht: (2025)
A Review of Electromagnetic Elimination Methods for low-field portable MRI scanner
von: Bian, Wanyu, et al.
Veröffentlicht: (2024)
von: Bian, Wanyu, et al.
Veröffentlicht: (2024)
ToolVQA: A Dataset for Multi-step Reasoning VQA with External Tools
von: Yin, Shaofeng, et al.
Veröffentlicht: (2025)
von: Yin, Shaofeng, et al.
Veröffentlicht: (2025)
Safe Reinforcement Learning Beyond Baseline Control: A Hierarchical Framework for Space Triangle Tethered Formation System
von: Tao, Xinyi, et al.
Veröffentlicht: (2026)
von: Tao, Xinyi, et al.
Veröffentlicht: (2026)
CoTBox-TTT: Grounding Medical VQA with Visual Chain-of-Thought Boxes During Test-time Training
von: Qian, Jiahe, et al.
Veröffentlicht: (2025)
von: Qian, Jiahe, et al.
Veröffentlicht: (2025)
Visual Grounding Methods for VQA are Working for the Wrong Reasons!
von: Shrestha, Robik, et al.
Veröffentlicht: (2020)
von: Shrestha, Robik, et al.
Veröffentlicht: (2020)
Uncovering the Full Potential of Visual Grounding Methods in VQA
von: Reich, Daniel, et al.
Veröffentlicht: (2024)
von: Reich, Daniel, et al.
Veröffentlicht: (2024)
DVLTA-VQA: Decoupled Vision-Language Modeling with Text-Guided Adaptation for Blind Video Quality Assessment
von: Yu, Li, et al.
Veröffentlicht: (2025)
von: Yu, Li, et al.
Veröffentlicht: (2025)
Abduction of Domain Relationships from Data for VQA
von: Chowdhury, Al Mehdi Saadat, et al.
Veröffentlicht: (2025)
von: Chowdhury, Al Mehdi Saadat, et al.
Veröffentlicht: (2025)
Large Language models for Time Series Analysis: Techniques, Applications, and Challenges
von: Shi, Feifei, et al.
Veröffentlicht: (2025)
von: Shi, Feifei, et al.
Veröffentlicht: (2025)
DDFP: Data-dependent Frequency Prompt for Source Free Domain Adaptation of Medical Image Segmentation
von: Yin, Siqi, et al.
Veröffentlicht: (2025)
von: Yin, Siqi, et al.
Veröffentlicht: (2025)
SemiDAViL: Semi-supervised Domain Adaptation with Vision-Language Guidance for Semantic Segmentation
von: Basak, Hritam, et al.
Veröffentlicht: (2025)
von: Basak, Hritam, et al.
Veröffentlicht: (2025)
Few-Shot Adaptation of Grounding DINO for Agricultural Domain
von: Singh, Rajhans, et al.
Veröffentlicht: (2025)
von: Singh, Rajhans, et al.
Veröffentlicht: (2025)
Multi-Task Learning for Visually Grounded Reasoning in Gastrointestinal VQA
von: Safwan, Itbaan, et al.
Veröffentlicht: (2025)
von: Safwan, Itbaan, et al.
Veröffentlicht: (2025)
Entropy Minimization without Model Collapse: Mitigating Prediction Bias in Medical Imaging
von: Nielen, Tim, et al.
Veröffentlicht: (2026)
von: Nielen, Tim, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
von: Le, Chenqian, et al.
Veröffentlicht: (2025) -
Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering
von: Chen, Xupeng, et al.
Veröffentlicht: (2026) -
Harnessing Earnings Reports for Stock Predictions: A QLoRA-Enhanced LLM Approach
von: Ni, Haowei, et al.
Veröffentlicht: (2024) -
Enhancing Exchange Rate Forecasting with Explainable Deep Learning Models
von: Meng, Shuchen, et al.
Veröffentlicht: (2024) -
Time Series Modeling for Heart Rate Prediction: From ARIMA to Transformers
von: Ni, Haowei, et al.
Veröffentlicht: (2024)