Guardado en:
| Autores principales: | Pan, Dong, Li, Bingtao, Zheng, Yongsheng, Ma, Jiren, Fei, Victor |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.08019 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Comprehensive Survey of Mixture-of-Experts: Algorithms, Theory, and Applications
por: Mu, Siyuan, et al.
Publicado: (2025)
por: Mu, Siyuan, et al.
Publicado: (2025)
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025)
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025)
Exploring Expert Specialization through Unsupervised Training in Sparse Mixture of Experts
por: Nikolic, Strahinja, et al.
Publicado: (2025)
por: Nikolic, Strahinja, et al.
Publicado: (2025)
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
por: Panda, Ashwinee, et al.
Publicado: (2025)
por: Panda, Ashwinee, et al.
Publicado: (2025)
SMILE: Zero-Shot Sparse Mixture of Low-Rank Experts Construction From Pre-Trained Foundation Models
por: Tang, Anke, et al.
Publicado: (2024)
por: Tang, Anke, et al.
Publicado: (2024)
PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference
por: Zhao, Yushu, et al.
Publicado: (2025)
por: Zhao, Yushu, et al.
Publicado: (2025)
GRIP: Algorithm-Agnostic Machine Unlearning for Mixture-of-Experts via Geometric Router Constraints
por: Zhu, Andy, et al.
Publicado: (2026)
por: Zhu, Andy, et al.
Publicado: (2026)
Eradicating Negative Transfer in Multi-Physics Foundation Models via Sparse Mixture-of-Experts Routing
por: Sharma, Ellwil, et al.
Publicado: (2026)
por: Sharma, Ellwil, et al.
Publicado: (2026)
Soft-to-Hard Routing in Sparse Mixture-of-Experts Models
por: Rastegar, Reza
Publicado: (2026)
por: Rastegar, Reza
Publicado: (2026)
Sparse Models, Sparse Safety: Unsafe Routes in Mixture-of-Experts LLMs
por: Jiang, Yukun, et al.
Publicado: (2026)
por: Jiang, Yukun, et al.
Publicado: (2026)
Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models
por: Pan, Bowen, et al.
Publicado: (2024)
por: Pan, Bowen, et al.
Publicado: (2024)
From Sparse to Soft Mixtures of Experts
por: Puigcerver, Joan, et al.
Publicado: (2023)
por: Puigcerver, Joan, et al.
Publicado: (2023)
HodgeCover: Higher-Order Topological Coverage Drives Compression of Sparse Mixture-of-Experts
por: Zhong, Tao, et al.
Publicado: (2026)
por: Zhong, Tao, et al.
Publicado: (2026)
Revisiting Incremental Stochastic Majorization-Minimization Algorithms with Applications to Mixture of Experts
por: Tran, TrungKhang, et al.
Publicado: (2026)
por: Tran, TrungKhang, et al.
Publicado: (2026)
dFLMoE: Decentralized Federated Learning via Mixture of Experts for Medical Data Analysis
por: Xie, Luyuan, et al.
Publicado: (2025)
por: Xie, Luyuan, et al.
Publicado: (2025)
Unified Class and Domain Incremental Learning with Mixture of Experts for Indoor Localization
por: Singampalli, Akhil, et al.
Publicado: (2025)
por: Singampalli, Akhil, et al.
Publicado: (2025)
A Stronger Mixture of Low-Rank Experts for Fine-Tuning Foundation Models
por: Sun, Mengyang, et al.
Publicado: (2025)
por: Sun, Mengyang, et al.
Publicado: (2025)
UniPool: A Globally Shared Expert Pool for Mixture-of-Experts
por: Huang, Minbin, et al.
Publicado: (2026)
por: Huang, Minbin, et al.
Publicado: (2026)
The Illusion of Specialization: Unveiling the Domain-Invariant "Standing Committee" in Mixture-of-Experts Models
por: Wang, Yan, et al.
Publicado: (2026)
por: Wang, Yan, et al.
Publicado: (2026)
Mixture-of-Experts Meets In-Context Reinforcement Learning
por: Wu, Wenhao, et al.
Publicado: (2025)
por: Wu, Wenhao, et al.
Publicado: (2025)
Mixture of Raytraced Experts
por: Perin, Andrea, et al.
Publicado: (2025)
por: Perin, Andrea, et al.
Publicado: (2025)
Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
por: Shi, Xiaoming, et al.
Publicado: (2024)
por: Shi, Xiaoming, et al.
Publicado: (2024)
Wavelet Mixture of Experts for Time Series Forecasting
por: Zhou, Zheng, et al.
Publicado: (2025)
por: Zhou, Zheng, et al.
Publicado: (2025)
Drop-Upcycling: Training Sparse Mixture of Experts with Partial Re-initialization
por: Nakamura, Taishi, et al.
Publicado: (2025)
por: Nakamura, Taishi, et al.
Publicado: (2025)
Routing-Free Mixture-of-Experts
por: Liu, Yilun, et al.
Publicado: (2026)
por: Liu, Yilun, et al.
Publicado: (2026)
Mixture of Experts in a Mixture of RL settings
por: Willi, Timon, et al.
Publicado: (2024)
por: Willi, Timon, et al.
Publicado: (2024)
Speculating Experts Accelerates Inference for Mixture-of-Experts
por: Madan, Vivan, et al.
Publicado: (2026)
por: Madan, Vivan, et al.
Publicado: (2026)
MoFE-Time: Mixture of Frequency Domain Experts for Time-Series Forecasting Models
por: Liu, Yiwen, et al.
Publicado: (2025)
por: Liu, Yiwen, et al.
Publicado: (2025)
MIDG: Mixture of Invariant Experts with knowledge injection for Domain Generalization in Multimodal Sentiment Analysis
por: Li, Yangle, et al.
Publicado: (2025)
por: Li, Yangle, et al.
Publicado: (2025)
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
por: Wei, Jia, et al.
Publicado: (2026)
por: Wei, Jia, et al.
Publicado: (2026)
Integration of Mixture of Experts and Multimodal Generative AI in Internet of Vehicles: A Survey
por: Xu, Minrui, et al.
Publicado: (2024)
por: Xu, Minrui, et al.
Publicado: (2024)
Klotski: Efficient Mixture-of-Expert Inference via Expert-Aware Multi-Batch Pipeline
por: Fang, Zhiyuan, et al.
Publicado: (2025)
por: Fang, Zhiyuan, et al.
Publicado: (2025)
TT-LoRA MoE: Unifying Parameter-Efficient Fine-Tuning and Sparse Mixture-of-Experts
por: Kunwar, Pradip, et al.
Publicado: (2025)
por: Kunwar, Pradip, et al.
Publicado: (2025)
MC#: Mixture Compressor for Mixture-of-Experts Large Models
por: Huang, Wei, et al.
Publicado: (2025)
por: Huang, Wei, et al.
Publicado: (2025)
MoNDE: Mixture of Near-Data Experts for Large-Scale Sparse Models
por: Kim, Taehyun, et al.
Publicado: (2024)
por: Kim, Taehyun, et al.
Publicado: (2024)
MLPMoE: Zero-Shot Architectural Metamorphosis of Dense LLM MLPs into Static Mixture-of-Experts
por: Novikov, Ivan
Publicado: (2025)
por: Novikov, Ivan
Publicado: (2025)
Mixture of Diverse Size Experts
por: Sun, Manxi, et al.
Publicado: (2024)
por: Sun, Manxi, et al.
Publicado: (2024)
Sparsity and Superposition in Mixture of Experts
por: Chaudhari, Marmik, et al.
Publicado: (2025)
por: Chaudhari, Marmik, et al.
Publicado: (2025)
Mixture of Concept Bottleneck Experts
por: De Santis, Francesco, et al.
Publicado: (2026)
por: De Santis, Francesco, et al.
Publicado: (2026)
Mixture of A Million Experts
por: He, Xu Owen
Publicado: (2024)
por: He, Xu Owen
Publicado: (2024)
Ejemplares similares
-
A Comprehensive Survey of Mixture-of-Experts: Algorithms, Theory, and Applications
por: Mu, Siyuan, et al.
Publicado: (2025) -
Modeling Expert Interactions in Sparse Mixture of Experts via Graph Structures
por: Nguyen-Nhat, Minh-Khoi, et al.
Publicado: (2025) -
Exploring Expert Specialization through Unsupervised Training in Sparse Mixture of Experts
por: Nikolic, Strahinja, et al.
Publicado: (2025) -
Dense Backpropagation Improves Training for Sparse Mixture-of-Experts
por: Panda, Ashwinee, et al.
Publicado: (2025) -
SMILE: Zero-Shot Sparse Mixture of Low-Rank Experts Construction From Pre-Trained Foundation Models
por: Tang, Anke, et al.
Publicado: (2024)