Explaining and Mitigating the Modality Gap in Contrastive Multimodal Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Yaras, Can, Chen, Siyi, Wang, Peng, Qu, Qing |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Linearly Separable Features in Shallow Nonlinear Networks: Width Scales Polynomially with Intrinsic Data Dimension
por: Xu, Alec S., et al.
Publicado: (2025)
por: Xu, Alec S., et al.
Publicado: (2025)
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
por: Cai, Rui, et al.
Publicado: (2025)
por: Cai, Rui, et al.
Publicado: (2025)
Deep Multimodal Learning with Missing Modality: A Survey
por: Wu, Renjie, et al.
Publicado: (2024)
por: Wu, Renjie, et al.
Publicado: (2024)
Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP
por: Eslami, Sedigheh, et al.
Publicado: (2024)
por: Eslami, Sedigheh, et al.
Publicado: (2024)
Token Activation Map to Visually Explain Multimodal LLMs
por: Li, Yi, et al.
Publicado: (2025)
por: Li, Yi, et al.
Publicado: (2025)
Distilled Prompt Learning for Incomplete Multimodal Survival Prediction
por: Xu, Yingxue, et al.
Publicado: (2025)
por: Xu, Yingxue, et al.
Publicado: (2025)
ConVis: Contrastive Decoding with Hallucination Visualization for Mitigating Hallucinations in Multimodal Large Language Models
por: Park, Yeji, et al.
Publicado: (2024)
por: Park, Yeji, et al.
Publicado: (2024)
Understanding and Mitigating Human-Labelling Errors in Supervised Contrastive Learning
por: Long, Zijun, et al.
Publicado: (2024)
por: Long, Zijun, et al.
Publicado: (2024)
X-VORTEX: Spatio-Temporal Contrastive Learning for Wake Vortex Trajectory Forecasting
por: Qu, Zhan, et al.
Publicado: (2026)
por: Qu, Zhan, et al.
Publicado: (2026)
Understanding Deep Representation Learning via Layerwise Feature Compression and Discrimination
por: Wang, Peng, et al.
Publicado: (2023)
por: Wang, Peng, et al.
Publicado: (2023)
Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment
por: Chang, Kai-Po, et al.
Publicado: (2025)
por: Chang, Kai-Po, et al.
Publicado: (2025)
AFD: Mitigating Feature Gap for Adversarial Robustness by Feature Disentanglement
por: Zhou, Nuoyan, et al.
Publicado: (2024)
por: Zhou, Nuoyan, et al.
Publicado: (2024)
Contrastive Regularization over LoRA for Multimodal Biomedical Image Incremental Learning
por: Zhang, Haojie, et al.
Publicado: (2025)
por: Zhang, Haojie, et al.
Publicado: (2025)
Two-Phase Dynamics of Interactions Explains the Starting Point of a DNN Learning Over-Fitted Features
por: Zhang, Junpeng, et al.
Publicado: (2024)
por: Zhang, Junpeng, et al.
Publicado: (2024)
Structure-aware Contrastive Learning for Diagram Understanding of Multimodal Models
por: Sasaki, Hiroshi
Publicado: (2025)
por: Sasaki, Hiroshi
Publicado: (2025)
Robust Multimodal Learning via Cross-Modal Proxy Tokens
por: Reza, Md Kaykobad, et al.
Publicado: (2025)
por: Reza, Md Kaykobad, et al.
Publicado: (2025)
Robult: Leveraging Redundancy and Modality Specific Features for Robust Multimodal Learning
por: Nguyen, Duy A., et al.
Publicado: (2025)
por: Nguyen, Duy A., et al.
Publicado: (2025)
Efficient Backdoor Defense in Multimodal Contrastive Learning: A Token-Level Unlearning Method for Mitigating Threats
por: Liu, Kuanrong, et al.
Publicado: (2024)
por: Liu, Kuanrong, et al.
Publicado: (2024)
Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs
por: Jo, Yujin, et al.
Publicado: (2026)
por: Jo, Yujin, et al.
Publicado: (2026)
InfMasking: Unleashing Synergistic Information by Contrastive Multimodal Interactions
por: Wen, Liangjian, et al.
Publicado: (2025)
por: Wen, Liangjian, et al.
Publicado: (2025)
I0T: Embedding Standardization Method Towards Zero Modality Gap
por: An, Na Min, et al.
Publicado: (2024)
por: An, Na Min, et al.
Publicado: (2024)
Non-negative Contrastive Learning
por: Wang, Yifei, et al.
Publicado: (2024)
por: Wang, Yifei, et al.
Publicado: (2024)
GLC++: Source-Free Universal Domain Adaptation through Global-Local Clustering and Contrastive Affinity Learning
por: Qu, Sanqing, et al.
Publicado: (2024)
por: Qu, Sanqing, et al.
Publicado: (2024)
Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited Modalities
por: Saporta, Adriel, et al.
Publicado: (2024)
por: Saporta, Adriel, et al.
Publicado: (2024)
Towards General Modality Translation with Contrastive and Predictive Latent Diffusion Bridge
por: Berman, Nimrod, et al.
Publicado: (2025)
por: Berman, Nimrod, et al.
Publicado: (2025)
Cross the Gap: Exposing the Intra-modal Misalignment in CLIP via Modality Inversion
por: Mistretta, Marco, et al.
Publicado: (2025)
por: Mistretta, Marco, et al.
Publicado: (2025)
TinyAlign: Boosting Lightweight Vision-Language Models by Mitigating Modal Alignment Bottlenecks
por: Hu, Yuanze, et al.
Publicado: (2025)
por: Hu, Yuanze, et al.
Publicado: (2025)
Uncovering and Mitigating Transient Blindness in Multimodal Model Editing
por: Han, Xiaoqi, et al.
Publicado: (2025)
por: Han, Xiaoqi, et al.
Publicado: (2025)
Closing the Modality Gap for Mixed Modality Search
por: Li, Binxu, et al.
Publicado: (2025)
por: Li, Binxu, et al.
Publicado: (2025)
SMART: Semantic Matching Contrastive Learning for Partially View-Aligned Clustering
por: Peng, Liang, et al.
Publicado: (2025)
por: Peng, Liang, et al.
Publicado: (2025)
Adaptive Multi-head Contrastive Learning
por: Wang, Lei, et al.
Publicado: (2023)
por: Wang, Lei, et al.
Publicado: (2023)
CroMe: Multimodal Fake News Detection using Cross-Modal Tri-Transformer and Metric Learning
por: Choi, Eunjee, et al.
Publicado: (2025)
por: Choi, Eunjee, et al.
Publicado: (2025)
Point Cloud Understanding via Attention-Driven Contrastive Learning
por: Wang, Yi, et al.
Publicado: (2024)
por: Wang, Yi, et al.
Publicado: (2024)
Improving Multi-Label Contrastive Learning by Leveraging Label Distribution
por: Chen, Ning, et al.
Publicado: (2025)
por: Chen, Ning, et al.
Publicado: (2025)
Retrieval Visual Contrastive Decoding to Mitigate Object Hallucinations in Large Vision-Language Models
por: Lee, Jihoon, et al.
Publicado: (2025)
por: Lee, Jihoon, et al.
Publicado: (2025)
Do Generated Data Always Help Contrastive Learning?
por: Wang, Yifei, et al.
Publicado: (2024)
por: Wang, Yifei, et al.
Publicado: (2024)
Contrastive Factor Analysis
por: Duan, Zhibin, et al.
Publicado: (2024)
por: Duan, Zhibin, et al.
Publicado: (2024)
Explaining Recovery Trajectories of Older Adults Post Lower-Limb Fracture Using Modality-wise Multiview Clustering and Large Language Models
por: Khan, Shehroz S., et al.
Publicado: (2025)
por: Khan, Shehroz S., et al.
Publicado: (2025)
SimO Loss: Anchor-Free Contrastive Loss for Fine-Grained Supervised Contrastive Learning
por: Bouhsine, Taha, et al.
Publicado: (2024)
por: Bouhsine, Taha, et al.
Publicado: (2024)
Multimodal Pathway: Improve Transformers with Irrelevant Data from Other Modalities
por: Zhang, Yiyuan, et al.
Publicado: (2024)
por: Zhang, Yiyuan, et al.
Publicado: (2024)
Ejemplares similares
-
Linearly Separable Features in Shallow Nonlinear Networks: Width Scales Polynomially with Intrinsic Data Dimension
por: Xu, Alec S., et al.
Publicado: (2025) -
Diagnosing and Mitigating Modality Interference in Multimodal Large Language Models
por: Cai, Rui, et al.
Publicado: (2025) -
Deep Multimodal Learning with Missing Modality: A Survey
por: Wu, Renjie, et al.
Publicado: (2024) -
Mitigate the Gap: Investigating Approaches for Improving Cross-Modal Alignment in CLIP
por: Eslami, Sedigheh, et al.
Publicado: (2024) -
Token Activation Map to Visually Explain Multimodal LLMs
por: Li, Yi, et al.
Publicado: (2025)