DynaMoE: Dynamic Token-Level Expert Activation with Layer-Wise Adaptive Capacity for Mixture-of-Experts Neural Networks
Fuente:
arXiv
Enregistré dans:
| Auteur principal: | Gülmez, Gökdeniz |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Gabliteration: Adaptive Multi-Directional Neural Weight Modification for Selective Behavioral Alteration in Large Language Models
par: Gülmez, Gökdeniz
Publié: (2025)
par: Gülmez, Gökdeniz
Publié: (2025)
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
par: Jin, Peng, et autres
Publié: (2024)
par: Jin, Peng, et autres
Publié: (2024)
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
par: Liu, Baihui, et autres
Publié: (2026)
par: Liu, Baihui, et autres
Publié: (2026)
LightMoE: Reducing Mixture-of-Experts Redundancy through Expert Replacing
par: Hao, Jiawei, et autres
Publié: (2026)
par: Hao, Jiawei, et autres
Publié: (2026)
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
par: Wei, Jia, et autres
Publié: (2026)
par: Wei, Jia, et autres
Publié: (2026)
Dynamic Adaptive Shared Experts with Grouped Multi-Head Attention Mixture of Experts
par: Li, Cheng, et autres
Publié: (2025)
par: Li, Cheng, et autres
Publié: (2025)
HyperMoE: Towards Better Mixture of Experts via Transferring Among Experts
par: Zhao, Hao, et autres
Publié: (2024)
par: Zhao, Hao, et autres
Publié: (2024)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
par: Chen, Yuanteng, et autres
Publié: (2025)
par: Chen, Yuanteng, et autres
Publié: (2025)
BuddyMoE: Exploiting Expert Redundancy to Accelerate Memory-Constrained Mixture-of-Experts Inference
par: Wang, Yun, et autres
Publié: (2025)
par: Wang, Yun, et autres
Publié: (2025)
DA-MoE: Addressing Depth-Sensitivity in Graph-Level Analysis through Mixture of Experts
par: Yao, Zelin, et autres
Publié: (2024)
par: Yao, Zelin, et autres
Publié: (2024)
SDG-MoE: Signed Debate Graph Mixture-of-Experts
par: Kulibaba, Stepan, et autres
Publié: (2026)
par: Kulibaba, Stepan, et autres
Publié: (2026)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
par: Zhao, Heng, et autres
Publié: (2026)
par: Zhao, Heng, et autres
Publié: (2026)
Mixture of Experts (MoE): A Big Data Perspective
par: Gan, Wensheng, et autres
Publié: (2025)
par: Gan, Wensheng, et autres
Publié: (2025)
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
par: Yan, Jiaming, et autres
Publié: (2025)
par: Yan, Jiaming, et autres
Publié: (2025)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
par: Xie, Zhitian, et autres
Publié: (2024)
par: Xie, Zhitian, et autres
Publié: (2024)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
par: Zou, Will Y., et autres
Publié: (2025)
par: Zou, Will Y., et autres
Publié: (2025)
Mixture of Nested Experts: Adaptive Processing of Visual Tokens
par: Jain, Gagan, et autres
Publié: (2024)
par: Jain, Gagan, et autres
Publié: (2024)
MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
par: Yang, Cheng, et autres
Publié: (2024)
par: Yang, Cheng, et autres
Publié: (2024)
PWC-MoE: Privacy-Aware Wireless Collaborative Mixture of Experts
par: Su, Yang, et autres
Publié: (2025)
par: Su, Yang, et autres
Publié: (2025)
MobileMoE: Scaling On-Device Mixture of Experts
par: Chen, Yanbei, et autres
Publié: (2026)
par: Chen, Yanbei, et autres
Publié: (2026)
OrdMoE: Preference Alignment via Hierarchical Expert Group Ranking in Multimodal Mixture-of-Experts LLMs
par: Gao, Yuting, et autres
Publié: (2025)
par: Gao, Yuting, et autres
Publié: (2025)
Dynamic Expert Quantization for Scalable Mixture-of-Experts Inference
par: Chu, Kexin, et autres
Publié: (2025)
par: Chu, Kexin, et autres
Publié: (2025)
MP-MoE: Matrix Profile-Guided Mixture of Experts for Precipitation Forecasting
par: Tran, Huyen Ngoc, et autres
Publié: (2026)
par: Tran, Huyen Ngoc, et autres
Publié: (2026)
MoIN: Mixture of Introvert Experts to Upcycle an LLM
par: Tejankar, Ajinkya, et autres
Publié: (2024)
par: Tejankar, Ajinkya, et autres
Publié: (2024)
MoNTA: Accelerating Mixture-of-Experts Training with Network-Traffc-Aware Parallel Optimization
par: Guo, Jingming, et autres
Publié: (2024)
par: Guo, Jingming, et autres
Publié: (2024)
PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference
par: Zhao, Yushu, et autres
Publié: (2025)
par: Zhao, Yushu, et autres
Publié: (2025)
Graph-Conditioned Mixture of Graph Neural Network Experts for Traffic Forecasting
par: Ghaffari, Amirhossein, et autres
Publié: (2026)
par: Ghaffari, Amirhossein, et autres
Publié: (2026)
Geometric Regularization in Mixture-of-Experts: The Disconnect Between Weights and Activations
par: Kim, Hyunjun
Publié: (2026)
par: Kim, Hyunjun
Publié: (2026)
Speculating Experts Accelerates Inference for Mixture-of-Experts
par: Madan, Vivan, et autres
Publié: (2026)
par: Madan, Vivan, et autres
Publié: (2026)
Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns
par: Bambhaniya, Abhimanyu, et autres
Publié: (2026)
par: Bambhaniya, Abhimanyu, et autres
Publié: (2026)
MoE-DisCo:Low Economy Cost Training Mixture-of-Experts Models
par: Ye, Xin, et autres
Publié: (2026)
par: Ye, Xin, et autres
Publié: (2026)
LatentMoE: Toward Optimal Accuracy per FLOP and Parameter in Mixture of Experts
par: Elango, Venmugil, et autres
Publié: (2026)
par: Elango, Venmugil, et autres
Publié: (2026)
Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
par: Shi, Xiaoming, et autres
Publié: (2024)
par: Shi, Xiaoming, et autres
Publié: (2024)
Flex-MoE: Modeling Arbitrary Modality Combination via the Flexible Mixture-of-Experts
par: Yun, Sukwon, et autres
Publié: (2024)
par: Yun, Sukwon, et autres
Publié: (2024)
MoE-Health: A Mixture of Experts Framework for Robust Multimodal Healthcare Prediction
par: Wang, Xiaoyang, et autres
Publié: (2025)
par: Wang, Xiaoyang, et autres
Publié: (2025)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
par: Teo, Rachel S. Y., et autres
Publié: (2025)
par: Teo, Rachel S. Y., et autres
Publié: (2025)
MoEUT: Mixture-of-Experts Universal Transformers
par: Csordás, Róbert, et autres
Publié: (2024)
par: Csordás, Róbert, et autres
Publié: (2024)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
par: Gu, Naibin, et autres
Publié: (2025)
par: Gu, Naibin, et autres
Publié: (2025)
AT-MoE: Adaptive Task-planning Mixture of Experts via LoRA Approach
par: Li, Xurui, et autres
Publié: (2024)
par: Li, Xurui, et autres
Publié: (2024)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
par: Taniguchi, Rei, et autres
Publié: (2026)
par: Taniguchi, Rei, et autres
Publié: (2026)
Documents similaires
-
Gabliteration: Adaptive Multi-Directional Neural Weight Modification for Selective Behavioral Alteration in Large Language Models
par: Gülmez, Gökdeniz
Publié: (2025) -
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
par: Jin, Peng, et autres
Publié: (2024) -
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
par: Liu, Baihui, et autres
Publié: (2026) -
LightMoE: Reducing Mixture-of-Experts Redundancy through Expert Replacing
par: Hao, Jiawei, et autres
Publié: (2026) -
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
par: Wei, Jia, et autres
Publié: (2026)