On the Risk of Misleading Reports: Diagnosing Textual Biases in Multimodal Clinical AI
Fuente:
arXiv
Saved in:
| Main Authors: | Restrepo, David, Ktena, Ira, Vakalopoulou, Maria, Christodoulidis, Stergios, Ferrante, Enzo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Medical Context Distorts Decisions in Clinical Vision Language Models
by: Restrepo, David, et al.
Published: (2026)
by: Restrepo, David, et al.
Published: (2026)
Fairness and Robustness of CLIP-Based Models for Chest X-rays
by: Sourget, Théo, et al.
Published: (2025)
by: Sourget, Théo, et al.
Published: (2025)
Mask-HybridGNet: Graph-based segmentation with emergent anatomical correspondence from pixel-level supervision
by: Gaggion, Nicolás, et al.
Published: (2026)
by: Gaggion, Nicolás, et al.
Published: (2026)
PVeRA: Probabilistic Vector-Based Random Matrix Adaptation
by: Fillioux, Leo, et al.
Published: (2025)
by: Fillioux, Leo, et al.
Published: (2025)
ViG-Bias: Visually Grounded Bias Discovery and Mitigation
by: Marani, Badr-Eddine, et al.
Published: (2024)
by: Marani, Badr-Eddine, et al.
Published: (2024)
SGPMIL: Sparse Gaussian Process Multiple Instance Learning
by: Lolos, Andreas, et al.
Published: (2025)
by: Lolos, Andreas, et al.
Published: (2025)
CAPRMIL: Context-Aware Patch Representations for Multiple Instance Learning
by: Lolos, Andreas, et al.
Published: (2025)
by: Lolos, Andreas, et al.
Published: (2025)
BayesAdapter: enhanced uncertainty estimation in CLIP few-shot adaptation
by: Morales-Álvarez, Pablo, et al.
Published: (2024)
by: Morales-Álvarez, Pablo, et al.
Published: (2024)
Controllable Latent Space Augmentation for Digital Pathology
by: Boutaj, Sofiène, et al.
Published: (2025)
by: Boutaj, Sofiène, et al.
Published: (2025)
Multimodal Carotid Risk Stratification with Large Vision-Language Models: Benchmarking, Fine-Tuning, and Clinical Insights
by: Tsolissou, Daphne, et al.
Published: (2025)
by: Tsolissou, Daphne, et al.
Published: (2025)
OCTOPUS: Enhancing the Spatial-Awareness of Vision SSMs with Multi-Dimensional Scans and Traversal Selection
by: Mahatha, Kunal, et al.
Published: (2026)
by: Mahatha, Kunal, et al.
Published: (2026)
SoC: Semantic Orthogonal Calibration for Test-Time Prompt Tuning
by: Fillioux, Leo, et al.
Published: (2026)
by: Fillioux, Leo, et al.
Published: (2026)
Full Conformal Adaptation of Medical Vision-Language Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
Are foundation models for computer vision good conformal predictors?
by: Fillioux, Leo, et al.
Published: (2024)
by: Fillioux, Leo, et al.
Published: (2024)
Class Adaptive Conformal Training
by: Marani, Badr-Eddine, et al.
Published: (2026)
by: Marani, Badr-Eddine, et al.
Published: (2026)
THUNDER: Tile-level Histopathology image UNDERstanding benchmark
by: Marza, Pierre, et al.
Published: (2025)
by: Marza, Pierre, et al.
Published: (2025)
Information Maximization for Long-Tailed Semi-Supervised Domain Generalization
by: Fillioux, Leo, et al.
Published: (2026)
by: Fillioux, Leo, et al.
Published: (2026)
Seeing Through Deception: Uncovering Misleading Creator Intent in Multimodal News with Vision-Language Models
by: Wu, Jiaying, et al.
Published: (2025)
by: Wu, Jiaying, et al.
Published: (2025)
CDG-MAE: Learning Correspondences from Diffusion Generated Views
by: Belagali, Varun, et al.
Published: (2025)
by: Belagali, Varun, et al.
Published: (2025)
TRACE: Textual Relevance Augmentation and Contextual Encoding for Multimodal Hate Detection
by: Koushik, Girish A., et al.
Published: (2025)
by: Koushik, Girish A., et al.
Published: (2025)
CrossCheck-Bench: Diagnosing Compositional Failures in Multimodal Conflict Resolution
by: Tian, Baoliang, et al.
Published: (2025)
by: Tian, Baoliang, et al.
Published: (2025)
Open Challenges on Fairness of Artificial Intelligence in Medical Imaging Applications
by: Ferrante, Enzo, et al.
Published: (2024)
by: Ferrante, Enzo, et al.
Published: (2024)
Unleashing Spatial Reasoning in Multimodal Large Language Models via Textual Representation Guided Reasoning
by: Hua, Jiacheng, et al.
Published: (2026)
by: Hua, Jiacheng, et al.
Published: (2026)
Quantifying and Mitigating Unimodal Biases in Multimodal Large Language Models: A Causal Perspective
by: Chen, Meiqi, et al.
Published: (2024)
by: Chen, Meiqi, et al.
Published: (2024)
Detecting Offensive Memes with Social Biases in Singapore Context Using Multimodal Large Language Models
by: Yuxuan, Cao, et al.
Published: (2025)
by: Yuxuan, Cao, et al.
Published: (2025)
Octopus v3: Technical Report for On-device Sub-billion Multimodal AI Agent
by: Chen, Wei, et al.
Published: (2024)
by: Chen, Wei, et al.
Published: (2024)
VisText-Mosquito: A Unified Multimodal Dataset for Visual Detection, Segmentation, and Textual Explanation on Mosquito Breeding Sites
by: Islam, Md. Adnanul, et al.
Published: (2025)
by: Islam, Md. Adnanul, et al.
Published: (2025)
The Cost of Context: Mitigating Textual Bias in Multimodal Retrieval-Augmented Generation
by: Jung, Hoin, et al.
Published: (2026)
by: Jung, Hoin, et al.
Published: (2026)
How Does the Textual Information Affect the Retrieval of Multimodal In-Context Learning?
by: Luo, Yang, et al.
Published: (2024)
by: Luo, Yang, et al.
Published: (2024)
Checkup2Action: A Multimodal Clinical Check-up Report Dataset for Patient-Oriented Action Card Generation
by: Xiang, Sike, et al.
Published: (2026)
by: Xiang, Sike, et al.
Published: (2026)
Semantic Textual Similarity Assessment in Chest X-ray Reports Using a Domain-Specific Cosine-Based Metric
by: Picha, Sayeh Gholipour, et al.
Published: (2024)
by: Picha, Sayeh Gholipour, et al.
Published: (2024)
Not All Similarities Are Created Equal: Leveraging Data-Driven Biases to Inform GenAI Copyright Disputes
by: Hacohen, Uri, et al.
Published: (2024)
by: Hacohen, Uri, et al.
Published: (2024)
Systemic Biases in Sign Language AI Research: A Deaf-Led Call to Reevaluate Research Agendas
by: Desai, Aashaka, et al.
Published: (2024)
by: Desai, Aashaka, et al.
Published: (2024)
CLEAR: Character Unlearning in Textual and Visual Modalities
by: Dontsov, Alexey, et al.
Published: (2024)
by: Dontsov, Alexey, et al.
Published: (2024)
CHARTOM: A Visual Theory-of-Mind Benchmark for LLMs on Misleading Charts
by: Bharti, Shubham, et al.
Published: (2024)
by: Bharti, Shubham, et al.
Published: (2024)
Fool Me Once? Contrasting Textual and Visual Explanations in a Clinical Decision-Support Setting
by: Kayser, Maxime, et al.
Published: (2024)
by: Kayser, Maxime, et al.
Published: (2024)
Clinical Cognition Alignment for Gastrointestinal Diagnosis with Multimodal LLMs
by: Zheng, Huan, et al.
Published: (2026)
by: Zheng, Huan, et al.
Published: (2026)
Beyond the Textual: Generating Coherent Visual Options for MCQs
by: Wang, Wanqiang, et al.
Published: (2025)
by: Wang, Wanqiang, et al.
Published: (2025)
Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models
by: Gan, Woody Haosheng, et al.
Published: (2025)
by: Gan, Woody Haosheng, et al.
Published: (2025)
A Similarity Paradigm Through Textual Regularization Without Forgetting
by: Cui, Fangming, et al.
Published: (2025)
by: Cui, Fangming, et al.
Published: (2025)
Similar Items
-
Medical Context Distorts Decisions in Clinical Vision Language Models
by: Restrepo, David, et al.
Published: (2026) -
Fairness and Robustness of CLIP-Based Models for Chest X-rays
by: Sourget, Théo, et al.
Published: (2025) -
Mask-HybridGNet: Graph-based segmentation with emergent anatomical correspondence from pixel-level supervision
by: Gaggion, Nicolás, et al.
Published: (2026) -
PVeRA: Probabilistic Vector-Based Random Matrix Adaptation
by: Fillioux, Leo, et al.
Published: (2025) -
ViG-Bias: Visually Grounded Bias Discovery and Mitigation
by: Marani, Badr-Eddine, et al.
Published: (2024)