DPL: Decoupled Prototype Learning for Enhancing Robustness of Vision-Language Transformers to Missing Modalities
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Jueqing, Qi, Yuanyuan, Yang, Xiaohao, Niu, Shuaicheng, Ke, Fucai, Zhou, Shujie, Tan, Wei, Lin, Jionghao, Buntine, Wray, Rezatofighi, Hamid, Du, Lan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ALScope: A Unified Toolkit for Deep Active Learning
by: Wu, Chenkai, et al.
Published: (2025)
by: Wu, Chenkai, et al.
Published: (2025)
Navigating Conflicting Views: Harnessing Trust for Learning
by: Lu, Jueqing, et al.
Published: (2024)
by: Lu, Jueqing, et al.
Published: (2024)
Multi-Label Bayesian Active Learning with Inter-Label Relationships
by: Qi, Yuanyuan, et al.
Published: (2024)
by: Qi, Yuanyuan, et al.
Published: (2024)
LLM Reading Tea Leaves: Automatically Evaluating Topic Models with Large Language Models
by: Yang, Xiaohao, et al.
Published: (2024)
by: Yang, Xiaohao, et al.
Published: (2024)
ARIS: Agentic and Relationship Intelligence System for Social Robots
by: Datta, Stavya, et al.
Published: (2026)
by: Datta, Stavya, et al.
Published: (2026)
Harnessing the Power of Beta Scoring in Deep Active Learning for Multi-Label Text Classification
by: Tan, Wei, et al.
Published: (2024)
by: Tan, Wei, et al.
Published: (2024)
Neural Topic Modeling with Large Language Models in the Loop
by: Yang, Xiaohao, et al.
Published: (2024)
by: Yang, Xiaohao, et al.
Published: (2024)
Next Generation Active Learning: Mixture of LLMs in the Loop
by: Qi, Yuanyuan, et al.
Published: (2026)
by: Qi, Yuanyuan, et al.
Published: (2026)
Gradient-Guided Modality Decoupling for Missing-Modality Robustness
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
HYDRA: A Hyper Agent for Dynamic Compositional Visual Reasoning
by: Ke, Fucai, et al.
Published: (2024)
by: Ke, Fucai, et al.
Published: (2024)
Mini-BEHAVIOR-Gran: Revealing U-Shaped Effects of Instruction Granularity on Language-Guided Embodied Agents
by: Huang, Sukai, et al.
Published: (2026)
by: Huang, Sukai, et al.
Published: (2026)
VerifiAgent: a Unified Verification Agent in Language Model Reasoning
by: Han, Jiuzhou, et al.
Published: (2025)
by: Han, Jiuzhou, et al.
Published: (2025)
Towards Uncertainty-Aware Language Agent
by: Han, Jiuzhou, et al.
Published: (2024)
by: Han, Jiuzhou, et al.
Published: (2024)
Reward Engineering for Generating Semi-structured Explanation
by: Han, Jiuzhou, et al.
Published: (2023)
by: Han, Jiuzhou, et al.
Published: (2023)
Uncertainty-Based Methods for Automated Process Reward Data Construction and Output Aggregation in Mathematical Reasoning
by: Han, Jiuzhou, et al.
Published: (2025)
by: Han, Jiuzhou, et al.
Published: (2025)
Multimodal Federated Learning with Missing Modality via Prototype Mask and Contrast
by: Bao, Guangyin, et al.
Published: (2023)
by: Bao, Guangyin, et al.
Published: (2023)
Assessing the Sensitivity and Alignment of FOL Closeness Metrics
by: Thatikonda, Ramya Keerthy, et al.
Published: (2025)
by: Thatikonda, Ramya Keerthy, et al.
Published: (2025)
Logical Reasoning with Outcome Reward Models for Test-Time Scaling
by: Thatikonda, Ramya Keerthy, et al.
Published: (2025)
by: Thatikonda, Ramya Keerthy, et al.
Published: (2025)
DPL: Spatial-Conditioned Diffusion Prototype Enhancement for One-Shot Medical Segmentation
by: Gao, Ziyuan, et al.
Published: (2025)
by: Gao, Ziyuan, et al.
Published: (2025)
Scalable Transformer for High Dimensional Multivariate Time Series Forecasting
by: Zhou, Xin, et al.
Published: (2024)
by: Zhou, Xin, et al.
Published: (2024)
MTP: A Dataset for Multi-Modal Turning Points in Casual Conversations
by: Ho, Gia-Bao Dinh, et al.
Published: (2024)
by: Ho, Gia-Bao Dinh, et al.
Published: (2024)
A Multi-Modal Neuro-Symbolic Approach for Spatial Reasoning-Based Visual Grounding in Robotics
by: Jahangard, Simindokht, et al.
Published: (2025)
by: Jahangard, Simindokht, et al.
Published: (2025)
Divide-Conquer Transformer Learning for Predicting Electric Vehicle Charging Events Using Smart Meter Data
by: Ke, Fucai, et al.
Published: (2024)
by: Ke, Fucai, et al.
Published: (2024)
A Survey on Out-of-Distribution Evaluation of Neural NLP Models
by: Li, Xinzhe, et al.
Published: (2023)
by: Li, Xinzhe, et al.
Published: (2023)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
by: Han, Jiuzhou, et al.
Published: (2023)
by: Han, Jiuzhou, et al.
Published: (2023)
PRA-PoE: Robust Multimodal Alzheimer's Diagnosis with Arbitrary Missing Modalities
by: Yang, Guangqian, et al.
Published: (2026)
by: Yang, Guangqian, et al.
Published: (2026)
NAVER: A Neuro-Symbolic Compositional Automaton for Visual Grounding with Explicit Logic Reasoning
by: Cai, Zhixi, et al.
Published: (2025)
by: Cai, Zhixi, et al.
Published: (2025)
MATA: A Trainable Hierarchical Automaton System for Multi-Agent Visual Reasoning
by: Cai, Zhixi, et al.
Published: (2026)
by: Cai, Zhixi, et al.
Published: (2026)
Improving Symbolic Translation of Language Models for Logical Reasoning
by: Thatikonda, Ramya Keerthy, et al.
Published: (2026)
by: Thatikonda, Ramya Keerthy, et al.
Published: (2026)
Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification
by: Thatikonda, Ramya Keerthy, et al.
Published: (2024)
by: Thatikonda, Ramya Keerthy, et al.
Published: (2024)
OntoMedRec: Logically-Pretrained Model-Agnostic Ontology Encoders for Medication Recommendation
by: Tan, Weicong, et al.
Published: (2024)
by: Tan, Weicong, et al.
Published: (2024)
VIEW2SPACE: Studying Multi-View Visual Reasoning from Sparse Observations
by: Ke, Fucai, et al.
Published: (2026)
by: Ke, Fucai, et al.
Published: (2026)
Calibrated Multimodal Representation Learning with Missing Modalities
by: Liu, Xiaohao, et al.
Published: (2025)
by: Liu, Xiaohao, et al.
Published: (2025)
MuAP: Multi-step Adaptive Prompt Learning for Vision-Language Model with Missing Modality
by: Dai, Ruiting, et al.
Published: (2024)
by: Dai, Ruiting, et al.
Published: (2024)
OMG-Agent: Toward Robust Missing Modality Generation with Decoupled Coarse-to-Fine Agentic Workflows
by: Dai, Ruiting, et al.
Published: (2026)
by: Dai, Ruiting, et al.
Published: (2026)
dinov3.seg: Open-Vocabulary Semantic Segmentation with DINOv3
by: Dutta, Saikat, et al.
Published: (2026)
by: Dutta, Saikat, et al.
Published: (2026)
JRDB-Pose3D: A Multi-person 3D Human Pose and Shape Estimation Dataset for Robotics
by: Biswas, Sandika, et al.
Published: (2026)
by: Biswas, Sandika, et al.
Published: (2026)
Social-MAE: Social Masked Autoencoder for Multi-person Motion Representation Learning
by: Ehsanpour, Mahsa, et al.
Published: (2024)
by: Ehsanpour, Mahsa, et al.
Published: (2024)
DifFUSER: Diffusion Model for Robust Multi-Sensor Fusion in 3D Object Detection and BEV Segmentation
by: Le, Duy-Tho, et al.
Published: (2024)
by: Le, Duy-Tho, et al.
Published: (2024)
Cross-Modal Prototype based Multimodal Federated Learning under Severely Missing Modality
by: Le, Huy Q., et al.
Published: (2024)
by: Le, Huy Q., et al.
Published: (2024)
Similar Items
-
ALScope: A Unified Toolkit for Deep Active Learning
by: Wu, Chenkai, et al.
Published: (2025) -
Navigating Conflicting Views: Harnessing Trust for Learning
by: Lu, Jueqing, et al.
Published: (2024) -
Multi-Label Bayesian Active Learning with Inter-Label Relationships
by: Qi, Yuanyuan, et al.
Published: (2024) -
LLM Reading Tea Leaves: Automatically Evaluating Topic Models with Large Language Models
by: Yang, Xiaohao, et al.
Published: (2024) -
ARIS: Agentic and Relationship Intelligence System for Social Robots
by: Datta, Stavya, et al.
Published: (2026)