Differential-informed Sample Selection Accelerates Multimodal Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Zihua, Hong, Feng, Chen, Mengxi, Chen, Pengyi, Liu, Benyuan, Yao, Jiangchao, Zhang, Ya, Wang, Yanfeng |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mitigating Noisy Correspondence by Geometrical Structure Consistency Learning
by: Zhao, Zihua, et al.
Published: (2024)
by: Zhao, Zihua, et al.
Published: (2024)
Dual-granularity Sinkhorn Distillation for Enhanced Learning from Long-tailed Noisy Data
by: Hong, Feng, et al.
Published: (2025)
by: Hong, Feng, et al.
Published: (2025)
UniChest: Conquer-and-Divide Pre-training for Multi-Source Chest X-Ray Classification
by: Dai, Tianjie, et al.
Published: (2023)
by: Dai, Tianjie, et al.
Published: (2023)
Exploring Training on Heterogeneous Data with Mixture of Low-rank Adapters
by: Zhou, Yuhang, et al.
Published: (2024)
by: Zhou, Yuhang, et al.
Published: (2024)
Domain-Inspired Sharpness-Aware Minimization Under Domain Shifts
by: Zhang, Ruipeng, et al.
Published: (2024)
by: Zhang, Ruipeng, et al.
Published: (2024)
Learning to Instruct for Visual Instruction Tuning
by: Zhou, Zhihan, et al.
Published: (2025)
by: Zhou, Zhihan, et al.
Published: (2025)
Zero-shot Composed Text-Image Retrieval
by: Liu, Yikun, et al.
Published: (2023)
by: Liu, Yikun, et al.
Published: (2023)
LamRA: Large Multimodal Model as Your Advanced Retrieval Assistant
by: Liu, Yikun, et al.
Published: (2024)
by: Liu, Yikun, et al.
Published: (2024)
A Sanity Check on Composed Image Retrieval
by: Liu, Yikun, et al.
Published: (2026)
by: Liu, Yikun, et al.
Published: (2026)
CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning
by: Wang, Yiping, et al.
Published: (2024)
by: Wang, Yiping, et al.
Published: (2024)
Multi-Modal Prototypes for Open-World Semantic Segmentation
by: Yang, Yuhuan, et al.
Published: (2023)
by: Yang, Yuhuan, et al.
Published: (2023)
G4Seg: Generation for Inexact Segmentation Refinement with Diffusion Models
by: Zhang, Tianjiao, et al.
Published: (2025)
by: Zhang, Tianjiao, et al.
Published: (2025)
POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch
by: Liu, Yikun, et al.
Published: (2026)
by: Liu, Yikun, et al.
Published: (2026)
MoMa: Modulating Mamba for Adapting Image Foundation Models to Video Recognition
by: Yang, Yuhuan, et al.
Published: (2025)
by: Yang, Yuhuan, et al.
Published: (2025)
Reprogramming Distillation for Medical Foundation Models
by: Zhou, Yuhang, et al.
Published: (2024)
by: Zhou, Yuhang, et al.
Published: (2024)
Low-Rank Knowledge Decomposition for Medical Foundation Models
by: Zhou, Yuhang, et al.
Published: (2024)
by: Zhou, Yuhang, et al.
Published: (2024)
One-Step Diffusion Transformer for Controllable Real-World Image Super-Resolution
by: Fang, Yushun, et al.
Published: (2025)
by: Fang, Yushun, et al.
Published: (2025)
Test-Time Multimodal Backdoor Detection by Contrastive Prompting
by: Niu, Yuwei, et al.
Published: (2024)
by: Niu, Yuwei, et al.
Published: (2024)
LoRKD: Low-Rank Knowledge Decomposition for Medical Foundation Models
by: Li, Haolin, et al.
Published: (2024)
by: Li, Haolin, et al.
Published: (2024)
ReMamber: Referring Image Segmentation with Mamba Twister
by: Yang, Yuhuan, et al.
Published: (2024)
by: Yang, Yuhuan, et al.
Published: (2024)
Contrastive Learning for Multimodal Human Activity Recognition with Limited Labeled Data
by: Jing, Long, et al.
Published: (2026)
by: Jing, Long, et al.
Published: (2026)
$\mathbb{X}$-Sample Contrastive Loss: Improving Contrastive Learning with Sample Similarity Graphs
by: Sobal, Vlad, et al.
Published: (2024)
by: Sobal, Vlad, et al.
Published: (2024)
A Generalization Theory of Cross-Modality Distillation with Contrastive Learning
by: Lin, Hangyu, et al.
Published: (2024)
by: Lin, Hangyu, et al.
Published: (2024)
UrbanFusion: Stochastic Multimodal Fusion for Contrastive Learning of Robust Spatial Representations
by: Mühlematter, Dominik J., et al.
Published: (2025)
by: Mühlematter, Dominik J., et al.
Published: (2025)
MaxInfo: A Training-Free Key-Frame Selection Method Using Maximum Volume for Enhanced Video Understanding
by: Li, Pengyi, et al.
Published: (2025)
by: Li, Pengyi, et al.
Published: (2025)
Decouple before Align: Visual Disentanglement Enhances Prompt Tuning
by: Zhang, Fei, et al.
Published: (2025)
by: Zhang, Fei, et al.
Published: (2025)
Explaining and Mitigating the Modality Gap in Contrastive Multimodal Learning
by: Yaras, Can, et al.
Published: (2024)
by: Yaras, Can, et al.
Published: (2024)
Generalized Contrastive Learning for Universal Multimodal Retrieval
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
Deep Reprogramming Distillation for Medical Foundation Models
by: Du, Siyuan, et al.
Published: (2026)
by: Du, Siyuan, et al.
Published: (2026)
Confidence-aware Contrastive Learning for Selective Classification
by: Wu, Yu-Chang, et al.
Published: (2024)
by: Wu, Yu-Chang, et al.
Published: (2024)
CFDNet: A Generalizable Foggy Stereo Matching Network with Contrastive Feature Distillation
by: Liu, Zihua, et al.
Published: (2024)
by: Liu, Zihua, et al.
Published: (2024)
SaFiRe: Saccade-Fixation Reiteration with Mamba for Referring Image Segmentation
by: Mao, Zhenjie, et al.
Published: (2025)
by: Mao, Zhenjie, et al.
Published: (2025)
Connecting Domains and Contrasting Samples: A Ladder for Domain Generalization
by: Wei, Tianxin, et al.
Published: (2025)
by: Wei, Tianxin, et al.
Published: (2025)
Enhancing Semi-Supervised Learning via Representative and Diverse Sample Selection
by: Shao, Qian, et al.
Published: (2024)
by: Shao, Qian, et al.
Published: (2024)
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
by: Yu, Geng, et al.
Published: (2024)
by: Yu, Geng, et al.
Published: (2024)
Sample Selection via Contrastive Fragmentation for Noisy Label Regression
by: Kim, Chris Dongjoo, et al.
Published: (2025)
by: Kim, Chris Dongjoo, et al.
Published: (2025)
Contrast-Unity for Partially-Supervised Temporal Sentence Grounding
by: Wang, Haicheng, et al.
Published: (2025)
by: Wang, Haicheng, et al.
Published: (2025)
Accelerating MRI with Longitudinally-informed Latent Posterior Sampling
by: Urman, Yonatan, et al.
Published: (2024)
by: Urman, Yonatan, et al.
Published: (2024)
GenMask: Adapting DiT for Segmentation via Direct Mask Generation
by: Yang, Yuhuan, et al.
Published: (2026)
by: Yang, Yuhuan, et al.
Published: (2026)
SelectMix: Enhancing Label Noise Robustness through Targeted Sample Mixing
by: Liu, Qiuhao, et al.
Published: (2025)
by: Liu, Qiuhao, et al.
Published: (2025)
Similar Items
-
Mitigating Noisy Correspondence by Geometrical Structure Consistency Learning
by: Zhao, Zihua, et al.
Published: (2024) -
Dual-granularity Sinkhorn Distillation for Enhanced Learning from Long-tailed Noisy Data
by: Hong, Feng, et al.
Published: (2025) -
UniChest: Conquer-and-Divide Pre-training for Multi-Source Chest X-Ray Classification
by: Dai, Tianjie, et al.
Published: (2023) -
Exploring Training on Heterogeneous Data with Mixture of Low-rank Adapters
by: Zhou, Yuhang, et al.
Published: (2024) -
Domain-Inspired Sharpness-Aware Minimization Under Domain Shifts
by: Zhang, Ruipeng, et al.
Published: (2024)