Envisioning MedCLIP: A Deep Dive into Explainability for Medical Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Hashmi, Anees Ur Rehman, Mahapatra, Dwarikanath, Yaqub, Mohammad |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedMerge: Merging Models for Effective Transfer Learning to Medical Imaging Tasks
by: Almakky, Ibrahim, et al.
Published: (2024)
by: Almakky, Ibrahim, et al.
Published: (2024)
DynaMMo: Dynamic Model Merging for Efficient Class Incremental Learning for Medical Images
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
DiMPLe -- Disentangled Multi-Modal Prompt Learning: Enhancing Out-Of-Distribution Alignment with Invariant and Spurious Feature Separation
by: Rahman, Umaima, et al.
Published: (2025)
by: Rahman, Umaima, et al.
Published: (2025)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
On Enhancing Brain Tumor Segmentation Across Diverse Populations with Convolutional Neural Networks
by: Maani, Fadillah, et al.
Published: (2024)
by: Maani, Fadillah, et al.
Published: (2024)
Continual Learning in Medical Imaging: A Survey and Practical Analysis
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
by: Qazi, Mohammad Areeb, et al.
Published: (2024)
FissionFusion: Fast Geometric Generation and Hierarchical Souping for Medical Image Analysis
by: Sanjeev, Santosh, et al.
Published: (2024)
by: Sanjeev, Santosh, et al.
Published: (2024)
Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift
by: Rahman, Umaima, et al.
Published: (2025)
by: Rahman, Umaima, et al.
Published: (2025)
Advanced Tumor Segmentation in Medical Imaging: An Ensemble Approach for BraTS 2023 Adult Glioma and Pediatric Tumor Tasks
by: Maani, Fadillah, et al.
Published: (2024)
by: Maani, Fadillah, et al.
Published: (2024)
MedCLIP-SAMv2: Towards Universal Text-Driven Medical Image Segmentation
by: Koleilat, Taha, et al.
Published: (2024)
by: Koleilat, Taha, et al.
Published: (2024)
MedCLIP-SAM: Bridging Text and Image Towards Universal Medical Image Segmentation
by: Koleilat, Taha, et al.
Published: (2024)
by: Koleilat, Taha, et al.
Published: (2024)
Backdoor Attack on Unpaired Medical Image-Text Foundation Models: A Pilot Study on MedCLIP
by: Jin, Ruinan, et al.
Published: (2024)
by: Jin, Ruinan, et al.
Published: (2024)
Integrating MedCLIP and Cross-Modal Fusion for Automatic Radiology Report Generation
by: Han, Qianhao, et al.
Published: (2024)
by: Han, Qianhao, et al.
Published: (2024)
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
by: Rahman, Umaima, et al.
Published: (2024)
by: Rahman, Umaima, et al.
Published: (2024)
XReal: Realistic Anatomy and Pathology-Aware X-ray Generation via Controllable Diffusion Model
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024)
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024)
On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
Slide-Level Prompt Learning with Vision Language Models for Few-Shot Multiple Instance Learning in Histopathology
by: Tomar, Devavrat, et al.
Published: (2025)
by: Tomar, Devavrat, et al.
Published: (2025)
FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound
by: Alasmawi, Hussain, et al.
Published: (2025)
by: Alasmawi, Hussain, et al.
Published: (2025)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
Stride-Net: Fairness-Aware Disentangled Representation Learning for Chest X-Ray Diagnosis
by: Rashid, Darakshan, et al.
Published: (2026)
by: Rashid, Darakshan, et al.
Published: (2026)
MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration
by: Peng, Jiahui, et al.
Published: (2026)
by: Peng, Jiahui, et al.
Published: (2026)
Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models
by: Monon, Mashrafi, et al.
Published: (2026)
by: Monon, Mashrafi, et al.
Published: (2026)
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
by: Kolli, Govinda, et al.
Published: (2026)
by: Kolli, Govinda, et al.
Published: (2026)
MedNNS: Supernet-based Medical Task-Adaptive Neural Network Search
by: Mecharbat, Lotfi Abdelkrim, et al.
Published: (2025)
by: Mecharbat, Lotfi Abdelkrim, et al.
Published: (2025)
An Explainable Deep Neural Network with Frequency-Aware Channel and Spatial Refinement for Flood Prediction in Sustainable Cities
by: Dar, Shahid Shafi, et al.
Published: (2025)
by: Dar, Shahid Shafi, et al.
Published: (2025)
MedSAD-CLIP: Supervised CLIP with Token-Patch Cross-Attention for Medical Anomaly Detection and Segmentation
by: Tran, Thuy Truong, et al.
Published: (2026)
by: Tran, Thuy Truong, et al.
Published: (2026)
XAI-CLIP: ROI-Guided Perturbation Framework for Explainable Medical Image Segmentation in Multimodal Vision-Language Models
by: Alzubaidi, Thuraya, et al.
Published: (2026)
by: Alzubaidi, Thuraya, et al.
Published: (2026)
Automated Diabetic Screening via Anterior Segment Ocular Imaging: A Deep Learning and Explainable AI Approach
by: Maqsood, Hasaan, et al.
Published: (2026)
by: Maqsood, Hasaan, et al.
Published: (2026)
What's in the Image? A Deep-Dive into the Vision of Vision Language Models
by: Kaduri, Omri, et al.
Published: (2024)
by: Kaduri, Omri, et al.
Published: (2024)
MedPromptX: Grounded Multimodal Prompting for Chest X-ray Diagnosis
by: Shaaban, Mai A., et al.
Published: (2024)
by: Shaaban, Mai A., et al.
Published: (2024)
SilVar-Med: A Speech-Driven Visual Language Model for Explainable Abnormality Detection in Medical Imaging
by: Pham, Tan-Hanh, et al.
Published: (2025)
by: Pham, Tan-Hanh, et al.
Published: (2025)
In-Model Merging for Enhancing the Robustness of Medical Imaging Classification Models
by: Wang, Hu, et al.
Published: (2025)
by: Wang, Hu, et al.
Published: (2025)
Noise is an Efficient Learner for Zero-Shot Vision-Language Models
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
MAFM^3: Modular Adaptation of Foundation Models for Multi-Modal Medical AI
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
MultiMedEval: A Benchmark and a Toolkit for Evaluating Medical Vision-Language Models
by: Royer, Corentin, et al.
Published: (2024)
by: Royer, Corentin, et al.
Published: (2024)
Hulu-Med: A Transparent Generalist Model towards Holistic Medical Vision-Language Understanding
by: Jiang, Songtao, et al.
Published: (2025)
by: Jiang, Songtao, et al.
Published: (2025)
Med-R1: Reinforcement Learning for Generalizable Medical Reasoning in Vision-Language Models
by: Lai, Yuxiang, et al.
Published: (2025)
by: Lai, Yuxiang, et al.
Published: (2025)
MaskedCLIP: Bridging the Masked and CLIP Space for Semi-Supervised Medical Vision-Language Pre-training
by: Zhu, Lei, et al.
Published: (2025)
by: Zhu, Lei, et al.
Published: (2025)
MuGa-VTON: Multi-Garment Virtual Try-On via Diffusion Transformers with Prompt Customization
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
DuPLUS: Dual-Prompt Vision-Language Framework for Universal Medical Image Segmentation and Prognosis
by: Saeed, Numan, et al.
Published: (2025)
by: Saeed, Numan, et al.
Published: (2025)
Similar Items
-
MedMerge: Merging Models for Effective Transfer Learning to Medical Imaging Tasks
by: Almakky, Ibrahim, et al.
Published: (2024) -
DynaMMo: Dynamic Model Merging for Efficient Class Incremental Learning for Medical Images
by: Qazi, Mohammad Areeb, et al.
Published: (2024) -
DiMPLe -- Disentangled Multi-Modal Prompt Learning: Enhancing Out-Of-Distribution Alignment with Invariant and Spurious Feature Separation
by: Rahman, Umaima, et al.
Published: (2025) -
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
by: Imam, Raza, et al.
Published: (2025) -
On Enhancing Brain Tumor Segmentation Across Diverse Populations with Convolutional Neural Networks
by: Maani, Fadillah, et al.
Published: (2024)