CLIP-UP: A Simple and Efficient Mixture-of-Experts CLIP Training Recipe with Sparse Upcycling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Xinze, Chen, Chen, Yang, Yinfei, Chen, Hong-You, Zhang, Bowen, Pal, Aditya, Zhu, Xiangxin, Du, Xianzhi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VeCLIP: Improving CLIP Training via Visual-enriched Captions
von: Lai, Zhengfeng, et al.
Veröffentlicht: (2023)
von: Lai, Zhengfeng, et al.
Veröffentlicht: (2023)
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
von: Zhang, Jihai, et al.
Veröffentlicht: (2024)
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025)
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
von: Vardi, Ben, et al.
Veröffentlicht: (2025)
von: Vardi, Ben, et al.
Veröffentlicht: (2025)
Efficient Training of Diffusion Mixture-of-Experts Models: A Practical Recipe
von: Liu, Yahui, et al.
Veröffentlicht: (2025)
von: Liu, Yahui, et al.
Veröffentlicht: (2025)
WMoE-CLIP: Wavelet-Enhanced Mixture-of-Experts Prompt Learning for Zero-Shot Anomaly Detection
von: Chen, Peng, et al.
Veröffentlicht: (2026)
von: Chen, Peng, et al.
Veröffentlicht: (2026)
Meta CLIP 2: A Worldwide Scaling Recipe
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
von: Chuang, Yung-Sung, et al.
Veröffentlicht: (2025)
un$^2$CLIP: Improving CLIP's Visual Detail Capturing Ability via Inverting unCLIP
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
Contrastive Localized Language-Image Pre-Training
von: Chen, Hong-You, et al.
Veröffentlicht: (2024)
von: Chen, Hong-You, et al.
Veröffentlicht: (2024)
NeuCLIP: Efficient Large-Scale CLIP Training with Neural Normalizer Optimization
von: Wei, Xiyuan, et al.
Veröffentlicht: (2025)
von: Wei, Xiyuan, et al.
Veröffentlicht: (2025)
NeuroCLIP: Neuromorphic Data Understanding by CLIP and SNN
von: Guo, Yufei, et al.
Veröffentlicht: (2023)
von: Guo, Yufei, et al.
Veröffentlicht: (2023)
AD-CLIP: Adapting Domains in Prompt Space Using CLIP
von: Singha, Mainak, et al.
Veröffentlicht: (2023)
von: Singha, Mainak, et al.
Veröffentlicht: (2023)
FastCLIP: A Suite of Optimization Techniques to Accelerate CLIP Training with Limited Resources
von: Wei, Xiyuan, et al.
Veröffentlicht: (2024)
von: Wei, Xiyuan, et al.
Veröffentlicht: (2024)
CLIP-Map: Structured Matrix Mapping for Parameter-Efficient CLIP Compression
von: Zhang, Kangjie, et al.
Veröffentlicht: (2026)
von: Zhang, Kangjie, et al.
Veröffentlicht: (2026)
CLIP-RD: Relative Distillation for Efficient CLIP Knowledge Distillation
von: Chung, Jeannie, et al.
Veröffentlicht: (2026)
von: Chung, Jeannie, et al.
Veröffentlicht: (2026)
CuMo: Scaling Multimodal LLM with Co-Upcycled Mixture-of-Experts
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
von: Li, Jiachen, et al.
Veröffentlicht: (2024)
A Training-Free Framework for Open-Vocabulary Image Segmentation and Recognition with EfficientNet and CLIP
von: Dai, Ying, et al.
Veröffentlicht: (2025)
von: Dai, Ying, et al.
Veröffentlicht: (2025)
CLIP-CID: Efficient CLIP Distillation via Cluster-Instance Discrimination
von: Yang, Kaicheng, et al.
Veröffentlicht: (2024)
von: Yang, Kaicheng, et al.
Veröffentlicht: (2024)
LPT++: Efficient Training on Mixture of Long-tailed Experts
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
von: Teo, Rachel S. Y., et al.
Veröffentlicht: (2025)
CRoF: CLIP-based Robust Few-shot Learning on Noisy Labels
von: Deng, Shizhuo, et al.
Veröffentlicht: (2024)
von: Deng, Shizhuo, et al.
Veröffentlicht: (2024)
MapExpert: Online HD Map Construction with Simple and Efficient Sparse Map Element Expert
von: Zhang, Dapeng, et al.
Veröffentlicht: (2024)
von: Zhang, Dapeng, et al.
Veröffentlicht: (2024)
CLIP-VIS: Adapting CLIP for Open-Vocabulary Video Instance Segmentation
von: Zhu, Wenqi, et al.
Veröffentlicht: (2024)
von: Zhu, Wenqi, et al.
Veröffentlicht: (2024)
VTD-CLIP: Video-to-Text Discretization via Prompting CLIP
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
von: Zhu, Wencheng, et al.
Veröffentlicht: (2025)
ProxyCLIP: Proxy Attention Improves CLIP for Open-Vocabulary Segmentation
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
ClearCLIP: Decomposing CLIP Representations for Dense Vision-Language Inference
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
von: Lan, Mengcheng, et al.
Veröffentlicht: (2024)
CLIP the Divergence: Language-guided Unsupervised Domain Adaptation
von: Zhu, Jinjing, et al.
Veröffentlicht: (2024)
von: Zhu, Jinjing, et al.
Veröffentlicht: (2024)
Breaking the Limits of Open-Weight CLIP: An Optimization Framework for Self-supervised Fine-tuning of CLIP
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
von: Mehta, Anant, et al.
Veröffentlicht: (2026)
MV-CLIP: Multi-View CLIP for Zero-shot 3D Shape Recognition
von: Song, Dan, et al.
Veröffentlicht: (2023)
von: Song, Dan, et al.
Veröffentlicht: (2023)
TiC-CLIP: Continual Training of CLIP Models
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
von: Garg, Saurabh, et al.
Veröffentlicht: (2023)
CLIP-KD: An Empirical Study of CLIP Model Distillation
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
von: Yang, Chuanguang, et al.
Veröffentlicht: (2023)
OFF-CLIP: Improving Normal Detection Confidence in Radiology CLIP with Simple Off-Diagonal Term Auto-Adjustment
von: Park, Junhyun, et al.
Veröffentlicht: (2025)
von: Park, Junhyun, et al.
Veröffentlicht: (2025)
MoP-CLIP: A Mixture of Prompt-Tuned CLIP Models for Domain Incremental Learning
von: Nicolas, Julien, et al.
Veröffentlicht: (2023)
von: Nicolas, Julien, et al.
Veröffentlicht: (2023)
CLIPCleaner: Cleaning Noisy Labels with CLIP
von: Feng, Chen, et al.
Veröffentlicht: (2024)
von: Feng, Chen, et al.
Veröffentlicht: (2024)
MoCLIP-Lite: Efficient Video Recognition by Fusing CLIP with Motion Vectors
von: Huang, Binhua, et al.
Veröffentlicht: (2025)
von: Huang, Binhua, et al.
Veröffentlicht: (2025)
CLIP-SLA: Parameter-Efficient CLIP Adaptation for Continuous Sign Language Recognition
von: Alyami, Sarah, et al.
Veröffentlicht: (2025)
von: Alyami, Sarah, et al.
Veröffentlicht: (2025)
CardiacCLIP: Video-based CLIP Adaptation for LVEF Prediction in a Few-shot Manner
von: Du, Yao, et al.
Veröffentlicht: (2025)
von: Du, Yao, et al.
Veröffentlicht: (2025)
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
von: Cai, Yuliang, et al.
Veröffentlicht: (2025)
IsoCLIP: Decomposing CLIP Projectors for Efficient Intra-modal Alignment
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
von: Magistri, Simone, et al.
Veröffentlicht: (2026)
Autoregressive Video Generation beyond Next Frames Prediction
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
von: Ren, Sucheng, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VeCLIP: Improving CLIP Training via Visual-enriched Captions
von: Lai, Zhengfeng, et al.
Veröffentlicht: (2023) -
CLIP-MoE: Towards Building Mixture of Experts for CLIP with Diversified Multiplet Upcycling
von: Zhang, Jihai, et al.
Veröffentlicht: (2024) -
SuperCLIP: CLIP with Simple Classification Supervision
von: Zhao, Weiheng, et al.
Veröffentlicht: (2025) -
CLIP-UP: CLIP-Based Unanswerable Problem Detection for Visual Question Answering
von: Vardi, Ben, et al.
Veröffentlicht: (2025) -
Efficient Training of Diffusion Mixture-of-Experts Models: A Practical Recipe
von: Liu, Yahui, et al.
Veröffentlicht: (2025)