EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Jing, Linglin, Gao, Yuting, Wang, Zhigang, Lan, Wang, Tang, Yiwen, Wang, Wenhai, Zhang, Kaipeng, Guo, Qingpei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
OrdMoE: Preference Alignment via Hierarchical Expert Group Ranking in Multimodal Mixture-of-Experts LLMs
por: Gao, Yuting, et al.
Publicado: (2025)
por: Gao, Yuting, et al.
Publicado: (2025)
AnyExperts: On-Demand Expert Allocation for Multimodal Language Models with Mixture of Expert
por: Gao, Yuting, et al.
Publicado: (2025)
por: Gao, Yuting, et al.
Publicado: (2025)
FourierMoE: Fourier Mixture-of-Experts Adaptation of Large Language Models
por: Jiang, Juyong, et al.
Publicado: (2026)
por: Jiang, Juyong, et al.
Publicado: (2026)
LLaVA-CMoE: Towards Continual Mixture of Experts for Large Vision-Language Models
por: Zhao, Hengyuan, et al.
Publicado: (2025)
por: Zhao, Hengyuan, et al.
Publicado: (2025)
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
por: Huang, Yushi, et al.
Publicado: (2025)
por: Huang, Yushi, et al.
Publicado: (2025)
A Survey on Mixture of Experts in Large Language Models
por: Cai, Weilin, et al.
Publicado: (2024)
por: Cai, Weilin, et al.
Publicado: (2024)
Cluster-Driven Expert Pruning for Mixture-of-Experts Large Language Models
por: Guo, Hongcheng, et al.
Publicado: (2025)
por: Guo, Hongcheng, et al.
Publicado: (2025)
DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
por: Dai, Damai, et al.
Publicado: (2024)
por: Dai, Damai, et al.
Publicado: (2024)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
por: Feng, Yuchen, et al.
Publicado: (2025)
por: Feng, Yuchen, et al.
Publicado: (2025)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
por: Zhou, Hao, et al.
Publicado: (2024)
por: Zhou, Hao, et al.
Publicado: (2024)
BLR-MoE: Boosted Language-Routing Mixture of Experts for Domain-Robust Multilingual E2E ASR
por: Ma, Guodong, et al.
Publicado: (2025)
por: Ma, Guodong, et al.
Publicado: (2025)
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
por: Tang, Yehui, et al.
Publicado: (2025)
por: Tang, Yehui, et al.
Publicado: (2025)
$μ$-MoE: Test-Time Pruning as Micro-Grained Mixture-of-Experts
por: Koike-Akino, Toshiaki, et al.
Publicado: (2025)
por: Koike-Akino, Toshiaki, et al.
Publicado: (2025)
Unveiling Super Experts in Mixture-of-Experts Large Language Models
por: Su, Zunhai, et al.
Publicado: (2025)
por: Su, Zunhai, et al.
Publicado: (2025)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
por: Li, Yunxin, et al.
Publicado: (2024)
por: Li, Yunxin, et al.
Publicado: (2024)
MultiPL-MoE: Multi-Programming-Lingual Extension of Large Language Models through Hybrid Mixture-of-Experts
por: Wang, Qing, et al.
Publicado: (2025)
por: Wang, Qing, et al.
Publicado: (2025)
LEMoE: Advanced Mixture of Experts Adaptor for Lifelong Model Editing of Large Language Models
por: Wang, Renzhi, et al.
Publicado: (2024)
por: Wang, Renzhi, et al.
Publicado: (2024)
$\infty$-MoE: Generalizing Mixture of Experts to Infinite Experts
por: Takashiro, Shota, et al.
Publicado: (2026)
por: Takashiro, Shota, et al.
Publicado: (2026)
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling
por: Jiang, Fan, et al.
Publicado: (2026)
por: Jiang, Fan, et al.
Publicado: (2026)
ToMoE: Converting Dense Large Language Models to Mixture-of-Experts through Dynamic Structural Pruning
por: Gao, Shangqian, et al.
Publicado: (2025)
por: Gao, Shangqian, et al.
Publicado: (2025)
TiMoE: Time-Aware Mixture of Language Experts
por: Faro, Robin, et al.
Publicado: (2025)
por: Faro, Robin, et al.
Publicado: (2025)
MoBiLE: Efficient Mixture-of-Experts Inference on Consumer GPU with Mixture of Big Little Experts
por: Zhao, Yushu, et al.
Publicado: (2025)
por: Zhao, Yushu, et al.
Publicado: (2025)
HMoE: Heterogeneous Mixture of Experts for Language Modeling
por: Wang, An, et al.
Publicado: (2024)
por: Wang, An, et al.
Publicado: (2024)
Linear-MoE: Linear Sequence Modeling Meets Mixture-of-Experts
por: Sun, Weigao, et al.
Publicado: (2025)
por: Sun, Weigao, et al.
Publicado: (2025)
A Closer Look into Mixture-of-Experts in Large Language Models
por: Lo, Ka Man, et al.
Publicado: (2024)
por: Lo, Ka Man, et al.
Publicado: (2024)
MoDEM: Mixture of Domain Expert Models
por: Simonds, Toby, et al.
Publicado: (2024)
por: Simonds, Toby, et al.
Publicado: (2024)
MoE-Pruner: Pruning Mixture-of-Experts Large Language Model using the Hints from Its Router
por: Xie, Yanyue, et al.
Publicado: (2024)
por: Xie, Yanyue, et al.
Publicado: (2024)
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
por: Liu, Baihui, et al.
Publicado: (2026)
por: Liu, Baihui, et al.
Publicado: (2026)
Mixture of Lookup Experts
por: Jie, Shibo, et al.
Publicado: (2025)
por: Jie, Shibo, et al.
Publicado: (2025)
When Are Experts Misrouted? Counterfactual Routing Analysis in Mixture-of-Experts Language Models
por: Yoon, Youngsik, et al.
Publicado: (2026)
por: Yoon, Youngsik, et al.
Publicado: (2026)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
por: Gu, Naibin, et al.
Publicado: (2025)
por: Gu, Naibin, et al.
Publicado: (2025)
Bayesian Mixture of Experts For Large Language Models
por: Dialameh, Maryam, et al.
Publicado: (2025)
por: Dialameh, Maryam, et al.
Publicado: (2025)
SEER-MoE: Sparse Expert Efficiency through Regularization for Mixture-of-Experts
por: Muzio, Alexandre, et al.
Publicado: (2024)
por: Muzio, Alexandre, et al.
Publicado: (2024)
MH-MoE: Multi-Head Mixture-of-Experts
por: Huang, Shaohan, et al.
Publicado: (2024)
por: Huang, Shaohan, et al.
Publicado: (2024)
Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models
por: Wei, Tianwen, et al.
Publicado: (2024)
por: Wei, Tianwen, et al.
Publicado: (2024)
Mixture of Heterogeneous Grouped Experts for Language Modeling
por: Ma, Zhicheng, et al.
Publicado: (2026)
por: Ma, Zhicheng, et al.
Publicado: (2026)
Pre-Attention Expert Prediction and Prefetching for Mixture-of-Experts Large Language Models
por: Zhu, Shien, et al.
Publicado: (2025)
por: Zhu, Shien, et al.
Publicado: (2025)
MMoE: Enhancing Multimodal Models with Mixtures of Multimodal Interaction Experts
por: Yu, Haofei, et al.
Publicado: (2023)
por: Yu, Haofei, et al.
Publicado: (2023)
MEMoE: Enhancing Model Editing with Mixture of Experts Adaptors
por: Wang, Renzhi, et al.
Publicado: (2024)
por: Wang, Renzhi, et al.
Publicado: (2024)
Not All Experts are Equal: Efficient Expert Pruning and Skipping for Mixture-of-Experts Large Language Models
por: Lu, Xudong, et al.
Publicado: (2024)
por: Lu, Xudong, et al.
Publicado: (2024)
Ejemplares similares
-
OrdMoE: Preference Alignment via Hierarchical Expert Group Ranking in Multimodal Mixture-of-Experts LLMs
por: Gao, Yuting, et al.
Publicado: (2025) -
AnyExperts: On-Demand Expert Allocation for Multimodal Language Models with Mixture of Expert
por: Gao, Yuting, et al.
Publicado: (2025) -
FourierMoE: Fourier Mixture-of-Experts Adaptation of Large Language Models
por: Jiang, Juyong, et al.
Publicado: (2026) -
LLaVA-CMoE: Towards Continual Mixture of Experts for Large Vision-Language Models
por: Zhao, Hengyuan, et al.
Publicado: (2025) -
MoDES: Accelerating Mixture-of-Experts Multimodal Large Language Models via Dynamic Expert Skipping
por: Huang, Yushi, et al.
Publicado: (2025)