DiMPLe -- Disentangled Multi-Modal Prompt Learning: Enhancing Out-Of-Distribution Alignment with Invariant and Spurious Feature Separation
Fuente:
arXiv
Saved in:
| Main Authors: | Rahman, Umaima, Yaqub, Mohammad, Mahapatra, Dwarikanath |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift
by: Rahman, Umaima, et al.
Published: (2025)
by: Rahman, Umaima, et al.
Published: (2025)
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
by: Rahman, Umaima, et al.
Published: (2024)
by: Rahman, Umaima, et al.
Published: (2024)
Envisioning MedCLIP: A Deep Dive into Explainability for Medical Vision-Language Models
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024)
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024)
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
by: Imam, Raza, et al.
Published: (2025)
by: Imam, Raza, et al.
Published: (2025)
Stride-Net: Fairness-Aware Disentangled Representation Learning for Chest X-Ray Diagnosis
by: Rashid, Darakshan, et al.
Published: (2026)
by: Rashid, Darakshan, et al.
Published: (2026)
Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models
by: Monon, Mashrafi, et al.
Published: (2026)
by: Monon, Mashrafi, et al.
Published: (2026)
ConDiSR: Contrastive Disentanglement and Style Regularization for Single Domain Generalization
by: Matsun, Aleksandr, et al.
Published: (2024)
by: Matsun, Aleksandr, et al.
Published: (2024)
MuGa-VTON: Multi-Garment Virtual Try-On via Diffusion Transformers with Prompt Customization
by: Deria, Ankan, et al.
Published: (2025)
by: Deria, Ankan, et al.
Published: (2025)
Slide-Level Prompt Learning with Vision Language Models for Few-Shot Multiple Instance Learning in Histopathology
by: Tomar, Devavrat, et al.
Published: (2025)
by: Tomar, Devavrat, et al.
Published: (2025)
InvariantOODG: Learning Invariant Features of Point Clouds for Out-of-Distribution Generalization
by: Zhang, Zhimin, et al.
Published: (2024)
by: Zhang, Zhimin, et al.
Published: (2024)
Learning Generalizable Models via Disentangling Spurious and Enhancing Potential Correlations
by: Wang, Na, et al.
Published: (2024)
by: Wang, Na, et al.
Published: (2024)
TPA: Temporal Prompt Alignment for Fetal Congenital Heart Defect Classification
by: Taratynova, Darya, et al.
Published: (2025)
by: Taratynova, Darya, et al.
Published: (2025)
Causal Disentanglement and Cross-Modal Alignment for Enhanced Few-Shot Learning
by: Jiang, Tianjiao, et al.
Published: (2025)
by: Jiang, Tianjiao, et al.
Published: (2025)
Spurious-Aware Prototype Refinement for Reliable Out-of-Distribution Detection
by: Zohrabi, Reihaneh, et al.
Published: (2025)
by: Zohrabi, Reihaneh, et al.
Published: (2025)
Improving Out-of-Distribution Detection with Disentangled Foreground and Background Features
by: Ding, Choubo, et al.
Published: (2023)
by: Ding, Choubo, et al.
Published: (2023)
MAFM^3: Modular Adaptation of Foundation Models for Multi-Modal Medical AI
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
by: Qazi, Mohammad Areeb, et al.
Published: (2025)
DICS: Find Domain-Invariant and Class-Specific Features for Out-of-Distribution Generalization
by: Miao, Qiaowei, et al.
Published: (2024)
by: Miao, Qiaowei, et al.
Published: (2024)
XDT-CXR: Investigating Cross-Disease Transferability in Zero-Shot Binary Classification of Chest X-Rays
by: Rahman, Umaima, et al.
Published: (2024)
by: Rahman, Umaima, et al.
Published: (2024)
Bayesian Cross-Modal Alignment Learning for Few-Shot Out-of-Distribution Generalization
by: Zhu, Lin, et al.
Published: (2025)
by: Zhu, Lin, et al.
Published: (2025)
MedPromptX: Grounded Multimodal Prompting for Chest X-ray Diagnosis
by: Shaaban, Mai A., et al.
Published: (2024)
by: Shaaban, Mai A., et al.
Published: (2024)
LATA: Laplacian-Assisted Transductive Adaptation for Conformal Uncertainty in Medical VLMs
by: Bozorgtabar, Behzad, et al.
Published: (2026)
by: Bozorgtabar, Behzad, et al.
Published: (2026)
Learning Transferable Negative Prompts for Out-of-Distribution Detection
by: Li, Tianqi, et al.
Published: (2024)
by: Li, Tianqi, et al.
Published: (2024)
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching
by: Liu, Yepeng, et al.
Published: (2025)
by: Liu, Yepeng, et al.
Published: (2025)
DiLO: Disentangled Latent Optimization for Learning Shape and Deformation in Grouped Deforming 3D Objects
by: Uddin, Mostofa Rafid, et al.
Published: (2025)
by: Uddin, Mostofa Rafid, et al.
Published: (2025)
Pursuing Feature Separation based on Neural Collapse for Out-of-Distribution Detection
by: Wu, Yingwen, et al.
Published: (2024)
by: Wu, Yingwen, et al.
Published: (2024)
VGS-Decoding: Visual Grounding Score Guided Decoding for Hallucination Mitigation in Medical VLMs
by: Kolli, Govinda, et al.
Published: (2026)
by: Kolli, Govinda, et al.
Published: (2026)
Toward Robust Medical Fairness: Debiased Dual-Modal Alignment via Text-Guided Attribute-Disentangled Prompt Learning for Vision-Language Models
by: Xia, Yuexuan, et al.
Published: (2025)
by: Xia, Yuexuan, et al.
Published: (2025)
Meta-Learned Modality-Weighted Knowledge Distillation for Robust Multi-Modal Learning with Missing Data
by: Wang, Hu, et al.
Published: (2024)
by: Wang, Hu, et al.
Published: (2024)
FOCoOp: Enhancing Out-of-Distribution Robustness in Federated Prompt Learning for Vision-Language Models
by: Liao, Xinting, et al.
Published: (2025)
by: Liao, Xinting, et al.
Published: (2025)
Class Is Invariant to Context and Vice Versa: On Learning Invariance for Out-Of-Distribution Generalization
by: Qi, Jiaxin, et al.
Published: (2022)
by: Qi, Jiaxin, et al.
Published: (2022)
Boost UAV-based Ojbect Detection via Scale-Invariant Feature Disentanglement and Adversarial Learning
by: Liu, Fan, et al.
Published: (2024)
by: Liu, Fan, et al.
Published: (2024)
PulmoFusion: Advancing Pulmonary Health with Efficient Multi-Modal Fusion
by: Sharshar, Ahmed, et al.
Published: (2025)
by: Sharshar, Ahmed, et al.
Published: (2025)
PEMMA: Parameter-Efficient Multi-Modal Adaptation for Medical Image Segmentation
by: Saadi, Nada, et al.
Published: (2024)
by: Saadi, Nada, et al.
Published: (2024)
Efficient Parameter Adaptation for Multi-Modal Medical Image Segmentation and Prognosis
by: Saeed, Numan, et al.
Published: (2025)
by: Saeed, Numan, et al.
Published: (2025)
FUSC: Fetal Ultrasound Semantic Clustering of Second Trimester Scans Using Deep Self-supervised Learning
by: Alasmawi, Hussain, et al.
Published: (2023)
by: Alasmawi, Hussain, et al.
Published: (2023)
Characterizing Continual Learning Scenarios and Strategies for Audio Analysis
by: Bhatt, Ruchi, et al.
Published: (2024)
by: Bhatt, Ruchi, et al.
Published: (2024)
Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction
by: Jing, Liu, et al.
Published: (2024)
by: Jing, Liu, et al.
Published: (2024)
GraphAlign: Enhancing Accurate Feature Alignment by Graph matching for Multi-Modal 3D Object Detection
by: Song, Ziying, et al.
Published: (2023)
by: Song, Ziying, et al.
Published: (2023)
CosmoCLIP: Generalizing Large Vision-Language Models for Astronomical Imaging
by: Imam, Raza, et al.
Published: (2024)
by: Imam, Raza, et al.
Published: (2024)
Modality-Agnostic Prompt Learning for Multi-Modal Camouflaged Object Detection
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Similar Items
-
Decoupling Clinical and Class-Agnostic Features for Reliable Few-Shot Adaptation under Shift
by: Rahman, Umaima, et al.
Published: (2025) -
Can language-guided unsupervised adaptation improve medical image classification using unpaired images and texts?
by: Rahman, Umaima, et al.
Published: (2024) -
Envisioning MedCLIP: A Deep Dive into Explainability for Medical Vision-Language Models
by: Hashmi, Anees Ur Rehman, et al.
Published: (2024) -
T3: Test-Time Model Merging in VLMs for Zero-Shot Medical Imaging Analysis
by: Imam, Raza, et al.
Published: (2025) -
Stride-Net: Fairness-Aware Disentangled Representation Learning for Chest X-Ray Diagnosis
by: Rashid, Darakshan, et al.
Published: (2026)