Expert Routing for Communication-Efficient MoE via Finite Expert Banks
Fuente:
arXiv
Saved in:
| Main Authors: | Salehi, Mohammad Reza Deylam, Khalesi, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mixture-of-Experts under Finite-Rate Gating: Communication--Generalization Trade-offs
by: Khalesi, Ali, et al.
Published: (2026)
by: Khalesi, Ali, et al.
Published: (2026)
Sparse In-Network Learning via Shortest-Path Backpropagation and Finite-Rate Gating
by: Salehi, Mohammad Reza Deylam
Published: (2026)
by: Salehi, Mohammad Reza Deylam
Published: (2026)
Typical Solutions of Multi-User Linearly-Decomposable Distributed Computing
by: Khalesi, Ali, et al.
Published: (2025)
by: Khalesi, Ali, et al.
Published: (2025)
Distributed Compression for Computation and Bounds on the Optimal Rate
by: Salehi, Mohammad Reza Deylam, et al.
Published: (2025)
by: Salehi, Mohammad Reza Deylam, et al.
Published: (2025)
Route Experts by Sequence, not by Token
by: Wen, Tiansheng, et al.
Published: (2025)
by: Wen, Tiansheng, et al.
Published: (2025)
Non-Linear Function Computation Broadcast
by: Salehi, Mohammad Reza Deylam, et al.
Published: (2025)
by: Salehi, Mohammad Reza Deylam, et al.
Published: (2025)
Multi-Server Multi-Function Distributed Computation
by: Malak, Derya, et al.
Published: (2024)
by: Malak, Derya, et al.
Published: (2024)
Orchestrating Heterogeneous Experts: A Scalable MoE Framework with Anisotropy-Preserving Fusion
by: Liu, Ye, et al.
Published: (2025)
by: Liu, Ye, et al.
Published: (2025)
Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging
by: Li, Lujun, et al.
Published: (2025)
by: Li, Lujun, et al.
Published: (2025)
MoE Routing Testbed: Studying Expert Specialization and Routing Behavior at Small Scale
by: Falke, Tobias, et al.
Published: (2026)
by: Falke, Tobias, et al.
Published: (2026)
Fast MoE Inference via Predictive Prefetching and Expert Replication
by: Jyothish, Ankit, et al.
Published: (2026)
by: Jyothish, Ankit, et al.
Published: (2026)
Awakening Dormant Experts:Counterfactual Routing to Mitigate MoE Hallucinations
by: Hu, Wentao, et al.
Published: (2026)
by: Hu, Wentao, et al.
Published: (2026)
Fundamental limits of distributed covariance matrix estimation via a conditional strong data processing inequality
by: Rahmani, Mohammad Reza, et al.
Published: (2025)
by: Rahmani, Mohammad Reza, et al.
Published: (2025)
MergeMoE: Efficient Compression of MoE Models via Expert Output Merging
by: Miao, Ruijie, et al.
Published: (2025)
by: Miao, Ruijie, et al.
Published: (2025)
Steering MoE LLMs via Expert (De)Activation
by: Fayyaz, Mohsen, et al.
Published: (2025)
by: Fayyaz, Mohsen, et al.
Published: (2025)
MoE-Spec: Expert Budgeting for Efficient Speculative Decoding
by: McDanel, Bradley, et al.
Published: (2026)
by: McDanel, Bradley, et al.
Published: (2026)
MoE-Infinity: Efficient MoE Inference on Personal Machines with Sparsity-Aware Expert Cache
by: Xue, Leyang, et al.
Published: (2024)
by: Xue, Leyang, et al.
Published: (2024)
SiftMoE: Similarity-Aware Energy-Efficient Expert Selection for Wireless Distributed MoE Inference
by: Chen, Qian, et al.
Published: (2026)
by: Chen, Qian, et al.
Published: (2026)
Horseshoe Mixtures-of-Experts (HS-MoE)
by: Polson, Nick, et al.
Published: (2026)
by: Polson, Nick, et al.
Published: (2026)
$\infty$-MoE: Generalizing Mixture of Experts to Infinite Experts
by: Takashiro, Shota, et al.
Published: (2026)
by: Takashiro, Shota, et al.
Published: (2026)
Harder Tasks Need More Experts: Dynamic Routing in MoE Models
by: Huang, Quzhe, et al.
Published: (2024)
by: Huang, Quzhe, et al.
Published: (2024)
Agentic AI-Based Joint Computing and Networking via Mixture of Experts and Large Language Models
by: Reifert, Robert-Jeron, et al.
Published: (2026)
by: Reifert, Robert-Jeron, et al.
Published: (2026)
Polysemantic Experts, Monosemantic Paths: Routing as Control in MoEs
by: Ye, Charles, et al.
Published: (2026)
by: Ye, Charles, et al.
Published: (2026)
DA-MoE: Towards Dynamic Expert Allocation for Mixture-of-Experts Models
by: Aghdam, Maryam Akhavan, et al.
Published: (2024)
by: Aghdam, Maryam Akhavan, et al.
Published: (2024)
Wireless Broadcast Gossip for Decentralized Drone Swarms: Success Probability, Contraction, and Optimal Aloha
by: Khalesi, Ali
Published: (2026)
by: Khalesi, Ali
Published: (2026)
General Multi-User Distributed Computing: A Learning-Theoretic RKHS Framework for Generic Nonlinear Target Functions with Topology-Aware Risk Analysis
by: Khalesi, Ali
Published: (2025)
by: Khalesi, Ali
Published: (2025)
Leave It to the Experts: Detecting Knowledge Distillation via MoE Expert Signatures
by: Li, Pingzhi, et al.
Published: (2025)
by: Li, Pingzhi, et al.
Published: (2025)
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
by: Jin, Peng, et al.
Published: (2024)
by: Jin, Peng, et al.
Published: (2024)
WDMoE: Wireless Distributed Large Language Models with Mixture of Experts
by: Xue, Nan, et al.
Published: (2024)
by: Xue, Nan, et al.
Published: (2024)
MoE Lens -- An Expert Is All You Need
by: Chaudhari, Marmik, et al.
Published: (2026)
by: Chaudhari, Marmik, et al.
Published: (2026)
MoE Pathfinder: Trajectory-driven Expert Pruning
by: Yang, Xican, et al.
Published: (2025)
by: Yang, Xican, et al.
Published: (2025)
Learning-Augmented Perfectly Secure Collaborative Matrix Multiplication
by: He, Zixuan, et al.
Published: (2026)
by: He, Zixuan, et al.
Published: (2026)
MoBE: Mixture-of-Basis-Experts for Compressing MoE-based LLMs
by: Chen, Xiaodong, et al.
Published: (2025)
by: Chen, Xiaodong, et al.
Published: (2025)
SEER-MoE: Sparse Expert Efficiency through Regularization for Mixture-of-Experts
by: Muzio, Alexandre, et al.
Published: (2024)
by: Muzio, Alexandre, et al.
Published: (2024)
LAER-MoE: Load-Adaptive Expert Re-layout for Efficient Mixture-of-Experts Training
by: Liu, Xinyi, et al.
Published: (2026)
by: Liu, Xinyi, et al.
Published: (2026)
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
by: Liu, Baihui, et al.
Published: (2026)
by: Liu, Baihui, et al.
Published: (2026)
AdapMoE: Adaptive Sensitivity-based Expert Gating and Management for Efficient MoE Inference
by: Zhong, Shuzhang, et al.
Published: (2024)
by: Zhong, Shuzhang, et al.
Published: (2024)
Exploiting the Experts: Unauthorized Compression in MoE-LLMs
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
by: Neogi, Pinaki Prasad Guha, et al.
Published: (2025)
MoE-Prism: Disentangling Monolithic Experts for Elastic MoE Services via Model-System Co-Designs
by: Xia, Xinfeng, et al.
Published: (2025)
by: Xia, Xinfeng, et al.
Published: (2025)
FLEX-MoE: Federated Mixture-of-Experts with Load-balanced Expert Assignment for Edge Computing
by: Zhang, Boyang, et al.
Published: (2025)
by: Zhang, Boyang, et al.
Published: (2025)
Similar Items
-
Mixture-of-Experts under Finite-Rate Gating: Communication--Generalization Trade-offs
by: Khalesi, Ali, et al.
Published: (2026) -
Sparse In-Network Learning via Shortest-Path Backpropagation and Finite-Rate Gating
by: Salehi, Mohammad Reza Deylam
Published: (2026) -
Typical Solutions of Multi-User Linearly-Decomposable Distributed Computing
by: Khalesi, Ali, et al.
Published: (2025) -
Distributed Compression for Computation and Bounds on the Optimal Rate
by: Salehi, Mohammad Reza Deylam, et al.
Published: (2025) -
Route Experts by Sequence, not by Token
by: Wen, Tiansheng, et al.
Published: (2025)