Interpreting Biomedical VLMs on High-Imbalance Out-of-Distributions: An Insight into BiomedCLIP on Radiology
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sadman, Nafiz, Zulkernine, Farhana, Kwan, Benjamin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Multimodal Approach For Endoscopic VCE Image Classification Using BiomedCLIP-PubMedBERT
von: Ganapathy, Nagarajan, et al.
Veröffentlicht: (2024)
von: Ganapathy, Nagarajan, et al.
Veröffentlicht: (2024)
BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
von: Zhang, Sheng, et al.
Veröffentlicht: (2023)
von: Zhang, Sheng, et al.
Veröffentlicht: (2023)
UniBiomed: A Universal Foundation Model for Grounded Biomedical Image Interpretation
von: Wu, Linshan, et al.
Veröffentlicht: (2025)
von: Wu, Linshan, et al.
Veröffentlicht: (2025)
ASMa: Asymmetric Spatio-temporal Masking for Skeleton Action Representation Learning
von: Anand, Aman, et al.
Veröffentlicht: (2026)
von: Anand, Aman, et al.
Veröffentlicht: (2026)
Depth-Guided Self-Supervised Human Keypoint Detection via Cross-Modal Distillation
von: Anand, Aman, et al.
Veröffentlicht: (2024)
von: Anand, Aman, et al.
Veröffentlicht: (2024)
Rethinking Out-of-Distribution Detection on Imbalanced Data Distribution
von: Liu, Kai, et al.
Veröffentlicht: (2024)
von: Liu, Kai, et al.
Veröffentlicht: (2024)
BiomedCoOp: Learning to Prompt for Biomedical Vision-Language Models
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
von: Koleilat, Taha, et al.
Veröffentlicht: (2024)
BiomedXPro: Prompt Optimization for Explainable Diagnosis with Biomedical Vision Language Models
von: Silva, Kaushitha, et al.
Veröffentlicht: (2025)
von: Silva, Kaushitha, et al.
Veröffentlicht: (2025)
Biomed-DPT: Dual Modality Prompt Tuning for Biomedical Vision-Language Models
von: Peng, Wei, et al.
Veröffentlicht: (2025)
von: Peng, Wei, et al.
Veröffentlicht: (2025)
Interpreting CLIP: Insights on the Robustness to ImageNet Distribution Shifts
von: Crabbé, Jonathan, et al.
Veröffentlicht: (2023)
von: Crabbé, Jonathan, et al.
Veröffentlicht: (2023)
Recent Advances in Out-of-Distribution Detection with CLIP-Like Models: A Survey
von: Li, Chaohua, et al.
Veröffentlicht: (2025)
von: Li, Chaohua, et al.
Veröffentlicht: (2025)
OFF-CLIP: Improving Normal Detection Confidence in Radiology CLIP with Simple Off-Diagonal Term Auto-Adjustment
von: Park, Junhyun, et al.
Veröffentlicht: (2025)
von: Park, Junhyun, et al.
Veröffentlicht: (2025)
UniCrossAdapter: Multimodal Adaptation of CLIP for Radiology Report Generation
von: Chen, Yaxiong, et al.
Veröffentlicht: (2025)
von: Chen, Yaxiong, et al.
Veröffentlicht: (2025)
Out of Sight, Not Out of Context? Egocentric Spatial Reasoning in VLMs Across Disjoint Frames
von: Ravi, Sahithya, et al.
Veröffentlicht: (2025)
von: Ravi, Sahithya, et al.
Veröffentlicht: (2025)
Your Data Is Not Perfect: Towards Cross-Domain Out-of-Distribution Detection in Class-Imbalanced Data
von: Fang, Xiang, et al.
Veröffentlicht: (2024)
von: Fang, Xiang, et al.
Veröffentlicht: (2024)
BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs
von: Singha, Mainak, et al.
Veröffentlicht: (2026)
von: Singha, Mainak, et al.
Veröffentlicht: (2026)
Integrating MedCLIP and Cross-Modal Fusion for Automatic Radiology Report Generation
von: Han, Qianhao, et al.
Veröffentlicht: (2024)
von: Han, Qianhao, et al.
Veröffentlicht: (2024)
Evaluating Vision Language Models (VLMs) for Radiology: A Comprehensive Analysis
von: Li, Frank, et al.
Veröffentlicht: (2025)
von: Li, Frank, et al.
Veröffentlicht: (2025)
VLMs Guided Interpretable Decision Making for Autonomous Driving
von: Hu, Xin, et al.
Veröffentlicht: (2025)
von: Hu, Xin, et al.
Veröffentlicht: (2025)
vMFCoOp: Towards Equilibrium on a Unified Hyperspherical Manifold for Prompting Biomedical VLMs
von: Shao, Minye, et al.
Veröffentlicht: (2025)
von: Shao, Minye, et al.
Veröffentlicht: (2025)
MedConcept: Unsupervised Concept Discovery for Interpretability in Medical VLMs
von: Haque, Md Rakibul, et al.
Veröffentlicht: (2026)
von: Haque, Md Rakibul, et al.
Veröffentlicht: (2026)
Interpreting the Second-Order Effects of Neurons in CLIP
von: Gandelsman, Yossi, et al.
Veröffentlicht: (2024)
von: Gandelsman, Yossi, et al.
Veröffentlicht: (2024)
Interpret, prune and distill Donut : towards lightweight VLMs for VQA on document
von: Mansour, Adnan Ben, et al.
Veröffentlicht: (2025)
von: Mansour, Adnan Ben, et al.
Veröffentlicht: (2025)
CRRG-CLIP: Automatic Generation of Chest Radiology Reports and Classification of Chest Radiographs
von: Xu, Jianfei, et al.
Veröffentlicht: (2024)
von: Xu, Jianfei, et al.
Veröffentlicht: (2024)
Compressing Model with Few Class-Imbalance Samples: An Out-of-Distribution Expedition
von: Wu, Tian-Shuang, et al.
Veröffentlicht: (2025)
von: Wu, Tian-Shuang, et al.
Veröffentlicht: (2025)
TriAug: Out-of-Distribution Detection for Imbalanced Breast Lesion in Ultrasound
von: Ye, Yinyu, et al.
Veröffentlicht: (2024)
von: Ye, Yinyu, et al.
Veröffentlicht: (2024)
Explanation-Aware Learning for Enhanced Interpretability in Biomedical Imaging
von: Faruqui, Zubair, et al.
Veröffentlicht: (2026)
von: Faruqui, Zubair, et al.
Veröffentlicht: (2026)
Sparse CLIP: Co-Optimizing Interpretability and Performance in Contrastive Learning
von: Qin, Chuan, et al.
Veröffentlicht: (2026)
von: Qin, Chuan, et al.
Veröffentlicht: (2026)
Concept-Enhanced Multimodal RAG: Towards Interpretable and Accurate Radiology Report Generation
von: Salmè, Marco, et al.
Veröffentlicht: (2026)
von: Salmè, Marco, et al.
Veröffentlicht: (2026)
Do VLMs Have Bad Eyes? Diagnosing Compositional Failures via Mechanistic Interpretability
von: Aravindan, Ashwath Vaithinathan, et al.
Veröffentlicht: (2025)
von: Aravindan, Ashwath Vaithinathan, et al.
Veröffentlicht: (2025)
Similarity-as-Evidence: Calibrating Overconfident VLMs for Interpretable and Label-Efficient Medical Active Learning
von: Xie, Zhuofan, et al.
Veröffentlicht: (2026)
von: Xie, Zhuofan, et al.
Veröffentlicht: (2026)
Hold-One-Shot-Out (HOSO) for Validation-Free Few-Shot CLIP Adapters
von: Vorster, Chris, et al.
Veröffentlicht: (2026)
von: Vorster, Chris, et al.
Veröffentlicht: (2026)
Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
von: Yeo, Wei Jie, et al.
Veröffentlicht: (2025)
Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models
von: Morelli, Fabian, et al.
Veröffentlicht: (2026)
von: Morelli, Fabian, et al.
Veröffentlicht: (2026)
Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection
von: Oh, Ahyoung, et al.
Veröffentlicht: (2026)
von: Oh, Ahyoung, et al.
Veröffentlicht: (2026)
A Diver Attention Estimation Framework for Effective Underwater Human-Robot Interaction
von: Enan, Sadman Sakib, et al.
Veröffentlicht: (2022)
von: Enan, Sadman Sakib, et al.
Veröffentlicht: (2022)
RadCLIP: Enhancing Radiologic Image Analysis through Contrastive Language-Image Pre-training
von: Lu, Zhixiu, et al.
Veröffentlicht: (2024)
von: Lu, Zhixiu, et al.
Veröffentlicht: (2024)
Interpretable Bilingual Multimodal Large Language Model for Diverse Biomedical Tasks
von: Wang, Lehan, et al.
Veröffentlicht: (2024)
von: Wang, Lehan, et al.
Veröffentlicht: (2024)
Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?
von: Park, Simon, et al.
Veröffentlicht: (2025)
von: Park, Simon, et al.
Veröffentlicht: (2025)
Interpreting and Analysing CLIP's Zero-Shot Image Classification via Mutual Knowledge
von: Sammani, Fawaz, et al.
Veröffentlicht: (2024)
von: Sammani, Fawaz, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Multimodal Approach For Endoscopic VCE Image Classification Using BiomedCLIP-PubMedBERT
von: Ganapathy, Nagarajan, et al.
Veröffentlicht: (2024) -
BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs
von: Zhang, Sheng, et al.
Veröffentlicht: (2023) -
UniBiomed: A Universal Foundation Model for Grounded Biomedical Image Interpretation
von: Wu, Linshan, et al.
Veröffentlicht: (2025) -
ASMa: Asymmetric Spatio-temporal Masking for Skeleton Action Representation Learning
von: Anand, Aman, et al.
Veröffentlicht: (2026) -
Depth-Guided Self-Supervised Human Keypoint Detection via Cross-Modal Distillation
von: Anand, Aman, et al.
Veröffentlicht: (2024)