On the Comparison between Multi-modal and Single-modal Contrastive Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Huang, Wei, Han, Andi, Chen, Yongqiang, Cao, Yuan, Xu, Zhiqiang, Suzuki, Taiji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Provably Transformers Harness Multi-Concept Word Semantics for Efficient In-Context Learning
von: Bu, Dake, et al.
Veröffentlicht: (2024)
von: Bu, Dake, et al.
Veröffentlicht: (2024)
Generalization Bound of Gradient Flow through Training Trajectory and Data-dependent Kernel
von: Chen, Yilan, et al.
Veröffentlicht: (2025)
von: Chen, Yilan, et al.
Veröffentlicht: (2025)
Backdoor Attacks on Multi-modal Contrastive Learning
von: Kuniyilh, Simi D, et al.
Veröffentlicht: (2026)
von: Kuniyilh, Simi D, et al.
Veröffentlicht: (2026)
On the Role of Label Noise in the Feature Learning Process
von: Han, Andi, et al.
Veröffentlicht: (2025)
von: Han, Andi, et al.
Veröffentlicht: (2025)
Mamba Can Learn Low-Dimensional Targets In-Context via Test-Time Feature Learning
von: Oh, Junsoo, et al.
Veröffentlicht: (2025)
von: Oh, Junsoo, et al.
Veröffentlicht: (2025)
On the Optimization and Generalization of Two-layer Transformers with Sign Gradient Descent
von: Li, Bingrui, et al.
Veröffentlicht: (2024)
von: Li, Bingrui, et al.
Veröffentlicht: (2024)
How Does Label Noise Gradient Descent Improve Generalization in the Low SNR Regime?
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine
von: Huang, Wei, et al.
Veröffentlicht: (2026)
von: Huang, Wei, et al.
Veröffentlicht: (2026)
On the Learning Dynamics of Two-layer Linear Networks with Label Noise SGD
von: Zhang, Tongcheng, et al.
Veröffentlicht: (2026)
von: Zhang, Tongcheng, et al.
Veröffentlicht: (2026)
Quantifying the Optimization and Generalization Advantages of Graph Neural Networks Over Multilayer Perceptrons
von: Huang, Wei, et al.
Veröffentlicht: (2023)
von: Huang, Wei, et al.
Veröffentlicht: (2023)
On the Feature Learning in Diffusion Models
von: Han, Andi, et al.
Veröffentlicht: (2024)
von: Han, Andi, et al.
Veröffentlicht: (2024)
Understanding the Robustness of Multi-modal Contrastive Learning to Distribution Shift
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
von: Xue, Yihao, et al.
Veröffentlicht: (2023)
Provable Benefit of Curriculum in Transformer Tree-Reasoning Post-Training
von: Bu, Dake, et al.
Veröffentlicht: (2025)
von: Bu, Dake, et al.
Veröffentlicht: (2025)
In-Context Learning Is Provably Bayesian Inference: A Generalization Theory for Meta-Learning
von: Wakayama, Tomoya, et al.
Veröffentlicht: (2025)
von: Wakayama, Tomoya, et al.
Veröffentlicht: (2025)
Provably Neural Active Learning Succeeds via Prioritizing Perplexing Samples
von: Bu, Dake, et al.
Veröffentlicht: (2024)
von: Bu, Dake, et al.
Veröffentlicht: (2024)
Transformers Learn Nonlinear Features In Context: Nonconvex Mean-field Dynamics on the Attention Landscape
von: Kim, Juno, et al.
Veröffentlicht: (2024)
von: Kim, Juno, et al.
Veröffentlicht: (2024)
Improving Medical Multi-modal Contrastive Learning with Expert Annotations
von: Kumar, Yogesh, et al.
Veröffentlicht: (2024)
von: Kumar, Yogesh, et al.
Veröffentlicht: (2024)
DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models
von: Bu, Dake, et al.
Veröffentlicht: (2026)
von: Bu, Dake, et al.
Veröffentlicht: (2026)
Provable In-Context Vector Arithmetic via Retrieving Task Concepts
von: Bu, Dake, et al.
Veröffentlicht: (2025)
von: Bu, Dake, et al.
Veröffentlicht: (2025)
CCPL: Cross-modal Contrastive Protein Learning
von: Zheng, Jiangbin, et al.
Veröffentlicht: (2023)
von: Zheng, Jiangbin, et al.
Veröffentlicht: (2023)
Learnable Cross-modal Knowledge Distillation for Multi-modal Learning with Missing Modality
von: Wang, Hu, et al.
Veröffentlicht: (2023)
von: Wang, Hu, et al.
Veröffentlicht: (2023)
Multi-modal Dynamic Proxy Learning for Personalized Multiple Clustering
von: Xu, Jinfeng, et al.
Veröffentlicht: (2025)
von: Xu, Jinfeng, et al.
Veröffentlicht: (2025)
Unsupervised Multi-modal Feature Alignment for Time Series Representation Learning
von: Liang, Chen, et al.
Veröffentlicht: (2023)
von: Liang, Chen, et al.
Veröffentlicht: (2023)
Multi-modal Transfer Learning between Biological Foundation Models
von: Garau-Luis, Juan Jose, et al.
Veröffentlicht: (2024)
von: Garau-Luis, Juan Jose, et al.
Veröffentlicht: (2024)
State Space Models are Provably Comparable to Transformers in Dynamic Token Selection
von: Nishikawa, Naoki, et al.
Veröffentlicht: (2024)
von: Nishikawa, Naoki, et al.
Veröffentlicht: (2024)
Transformers Provably Solve Parity Efficiently with Chain of Thought
von: Kim, Juno, et al.
Veröffentlicht: (2024)
von: Kim, Juno, et al.
Veröffentlicht: (2024)
Mean-field Analysis on Two-layer Neural Networks from a Kernel Perspective
von: Takakura, Shokichi, et al.
Veröffentlicht: (2024)
von: Takakura, Shokichi, et al.
Veröffentlicht: (2024)
Deep Two-Way Matrix Reordering for Relational Data Analysis
von: Watanabe, Chihiro, et al.
Veröffentlicht: (2021)
von: Watanabe, Chihiro, et al.
Veröffentlicht: (2021)
The Mechanism of Weak-to-Strong Generalization: Feature Elicitation from Latent Knowledge
von: Awano, Ryoya, et al.
Veröffentlicht: (2026)
von: Awano, Ryoya, et al.
Veröffentlicht: (2026)
AutoLL: Automatic Linear Layout of Graphs based on Deep Neural Network
von: Watanabe, Chihiro, et al.
Veröffentlicht: (2021)
von: Watanabe, Chihiro, et al.
Veröffentlicht: (2021)
Test time training enhances in-context learning of nonlinear functions
von: Kuwataka, Kento, et al.
Veröffentlicht: (2025)
von: Kuwataka, Kento, et al.
Veröffentlicht: (2025)
Transformers as Measure-Theoretic Associative Memory: A Statistical Perspective and Minimax Optimality
von: Kawata, Ryotaro, et al.
Veröffentlicht: (2026)
von: Kawata, Ryotaro, et al.
Veröffentlicht: (2026)
Approximation and Estimation Ability of Transformers for Sequence-to-Sequence Functions with Infinite Dimensional Input
von: Takakura, Shokichi, et al.
Veröffentlicht: (2023)
von: Takakura, Shokichi, et al.
Veröffentlicht: (2023)
Post-Training as Reweighting: A Stochastic View of Reasoning Trajectories in Language Models
von: Bu, Dake, et al.
Veröffentlicht: (2025)
von: Bu, Dake, et al.
Veröffentlicht: (2025)
Skin Lesion Phenotyping via Nested Multi-modal Contrastive Learning
von: Christopoulos, Dionysis, et al.
Veröffentlicht: (2025)
von: Christopoulos, Dionysis, et al.
Veröffentlicht: (2025)
Multi-modal Causal Structure Learning and Root Cause Analysis
von: Zheng, Lecheng, et al.
Veröffentlicht: (2024)
von: Zheng, Lecheng, et al.
Veröffentlicht: (2024)
Balancing Multi-modal Sensor Learning via Multi-objective Optimization
von: Fernando, Heshan, et al.
Veröffentlicht: (2025)
von: Fernando, Heshan, et al.
Veröffentlicht: (2025)
ETSCL: An Evidence Theory-Based Supervised Contrastive Learning Framework for Multi-modal Glaucoma Grading
von: Yang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Yang, Zhiyuan, et al.
Veröffentlicht: (2024)
Towards Multi-modal Transformers in Federated Learning
von: Sun, Guangyu, et al.
Veröffentlicht: (2024)
von: Sun, Guangyu, et al.
Veröffentlicht: (2024)
Multi-modal Learning for WebAssembly Reverse Engineering
von: Huang, Hanxian, et al.
Veröffentlicht: (2024)
von: Huang, Hanxian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Provably Transformers Harness Multi-Concept Word Semantics for Efficient In-Context Learning
von: Bu, Dake, et al.
Veröffentlicht: (2024) -
Generalization Bound of Gradient Flow through Training Trajectory and Data-dependent Kernel
von: Chen, Yilan, et al.
Veröffentlicht: (2025) -
Backdoor Attacks on Multi-modal Contrastive Learning
von: Kuniyilh, Simi D, et al.
Veröffentlicht: (2026) -
On the Role of Label Noise in the Feature Learning Process
von: Han, Andi, et al.
Veröffentlicht: (2025) -
Mamba Can Learn Low-Dimensional Targets In-Context via Test-Time Feature Learning
von: Oh, Junsoo, et al.
Veröffentlicht: (2025)