DynaMoE: Dynamic Token-Level Expert Activation with Layer-Wise Adaptive Capacity for Mixture-of-Experts Neural Networks
Fuente:
arXiv
Guardado en:
| Autor principal: | Gülmez, Gökdeniz |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Gabliteration: Adaptive Multi-Directional Neural Weight Modification for Selective Behavioral Alteration in Large Language Models
por: Gülmez, Gökdeniz
Publicado: (2025)
por: Gülmez, Gökdeniz
Publicado: (2025)
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
por: Jin, Peng, et al.
Publicado: (2024)
por: Jin, Peng, et al.
Publicado: (2024)
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
por: Liu, Baihui, et al.
Publicado: (2026)
por: Liu, Baihui, et al.
Publicado: (2026)
LightMoE: Reducing Mixture-of-Experts Redundancy through Expert Replacing
por: Hao, Jiawei, et al.
Publicado: (2026)
por: Hao, Jiawei, et al.
Publicado: (2026)
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
por: Wei, Jia, et al.
Publicado: (2026)
por: Wei, Jia, et al.
Publicado: (2026)
Dynamic Adaptive Shared Experts with Grouped Multi-Head Attention Mixture of Experts
por: Li, Cheng, et al.
Publicado: (2025)
por: Li, Cheng, et al.
Publicado: (2025)
HyperMoE: Towards Better Mixture of Experts via Transferring Among Experts
por: Zhao, Hao, et al.
Publicado: (2024)
por: Zhao, Hao, et al.
Publicado: (2024)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
por: Chen, Yuanteng, et al.
Publicado: (2025)
por: Chen, Yuanteng, et al.
Publicado: (2025)
BuddyMoE: Exploiting Expert Redundancy to Accelerate Memory-Constrained Mixture-of-Experts Inference
por: Wang, Yun, et al.
Publicado: (2025)
por: Wang, Yun, et al.
Publicado: (2025)
DA-MoE: Addressing Depth-Sensitivity in Graph-Level Analysis through Mixture of Experts
por: Yao, Zelin, et al.
Publicado: (2024)
por: Yao, Zelin, et al.
Publicado: (2024)
SDG-MoE: Signed Debate Graph Mixture-of-Experts
por: Kulibaba, Stepan, et al.
Publicado: (2026)
por: Kulibaba, Stepan, et al.
Publicado: (2026)
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
por: Zhao, Heng, et al.
Publicado: (2026)
por: Zhao, Heng, et al.
Publicado: (2026)
Mixture of Experts (MoE): A Big Data Perspective
por: Gan, Wensheng, et al.
Publicado: (2025)
por: Gan, Wensheng, et al.
Publicado: (2025)
Accelerating Mixture-of-Expert Inference with Adaptive Expert Split Mechanism
por: Yan, Jiaming, et al.
Publicado: (2025)
por: Yan, Jiaming, et al.
Publicado: (2025)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
por: Xie, Zhitian, et al.
Publicado: (2024)
por: Xie, Zhitian, et al.
Publicado: (2024)
Neural Inhibition Improves Dynamic Routing and Mixture of Experts
por: Zou, Will Y., et al.
Publicado: (2025)
por: Zou, Will Y., et al.
Publicado: (2025)
Mixture of Nested Experts: Adaptive Processing of Visual Tokens
por: Jain, Gagan, et al.
Publicado: (2024)
por: Jain, Gagan, et al.
Publicado: (2024)
MoE-I$^2$: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
por: Yang, Cheng, et al.
Publicado: (2024)
por: Yang, Cheng, et al.
Publicado: (2024)
PWC-MoE: Privacy-Aware Wireless Collaborative Mixture of Experts
por: Su, Yang, et al.
Publicado: (2025)
por: Su, Yang, et al.
Publicado: (2025)
MobileMoE: Scaling On-Device Mixture of Experts
por: Chen, Yanbei, et al.
Publicado: (2026)
por: Chen, Yanbei, et al.
Publicado: (2026)
OrdMoE: Preference Alignment via Hierarchical Expert Group Ranking in Multimodal Mixture-of-Experts LLMs
por: Gao, Yuting, et al.
Publicado: (2025)
por: Gao, Yuting, et al.
Publicado: (2025)
Dynamic Expert Quantization for Scalable Mixture-of-Experts Inference
por: Chu, Kexin, et al.
Publicado: (2025)
por: Chu, Kexin, et al.
Publicado: (2025)
MP-MoE: Matrix Profile-Guided Mixture of Experts for Precipitation Forecasting
por: Tran, Huyen Ngoc, et al.
Publicado: (2026)
por: Tran, Huyen Ngoc, et al.
Publicado: (2026)
MoIN: Mixture of Introvert Experts to Upcycle an LLM
por: Tejankar, Ajinkya, et al.
Publicado: (2024)
por: Tejankar, Ajinkya, et al.
Publicado: (2024)
MoNTA: Accelerating Mixture-of-Experts Training with Network-Traffc-Aware Parallel Optimization
por: Guo, Jingming, et al.
Publicado: (2024)
por: Guo, Jingming, et al.
Publicado: (2024)
PuzzleMoE: Efficient Compression of Large Mixture-of-Experts Models via Sparse Expert Merging and Bit-packed inference
por: Zhao, Yushu, et al.
Publicado: (2025)
por: Zhao, Yushu, et al.
Publicado: (2025)
Graph-Conditioned Mixture of Graph Neural Network Experts for Traffic Forecasting
por: Ghaffari, Amirhossein, et al.
Publicado: (2026)
por: Ghaffari, Amirhossein, et al.
Publicado: (2026)
Geometric Regularization in Mixture-of-Experts: The Disconnect Between Weights and Activations
por: Kim, Hyunjun
Publicado: (2026)
por: Kim, Hyunjun
Publicado: (2026)
Speculating Experts Accelerates Inference for Mixture-of-Experts
por: Madan, Vivan, et al.
Publicado: (2026)
por: Madan, Vivan, et al.
Publicado: (2026)
Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns
por: Bambhaniya, Abhimanyu, et al.
Publicado: (2026)
por: Bambhaniya, Abhimanyu, et al.
Publicado: (2026)
MoE-DisCo:Low Economy Cost Training Mixture-of-Experts Models
por: Ye, Xin, et al.
Publicado: (2026)
por: Ye, Xin, et al.
Publicado: (2026)
LatentMoE: Toward Optimal Accuracy per FLOP and Parameter in Mixture of Experts
por: Elango, Venmugil, et al.
Publicado: (2026)
por: Elango, Venmugil, et al.
Publicado: (2026)
Time-MoE: Billion-Scale Time Series Foundation Models with Mixture of Experts
por: Shi, Xiaoming, et al.
Publicado: (2024)
por: Shi, Xiaoming, et al.
Publicado: (2024)
Flex-MoE: Modeling Arbitrary Modality Combination via the Flexible Mixture-of-Experts
por: Yun, Sukwon, et al.
Publicado: (2024)
por: Yun, Sukwon, et al.
Publicado: (2024)
MoE-Health: A Mixture of Experts Framework for Robust Multimodal Healthcare Prediction
por: Wang, Xiaoyang, et al.
Publicado: (2025)
por: Wang, Xiaoyang, et al.
Publicado: (2025)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
por: Teo, Rachel S. Y., et al.
Publicado: (2025)
por: Teo, Rachel S. Y., et al.
Publicado: (2025)
MoEUT: Mixture-of-Experts Universal Transformers
por: Csordás, Róbert, et al.
Publicado: (2024)
por: Csordás, Róbert, et al.
Publicado: (2024)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
por: Gu, Naibin, et al.
Publicado: (2025)
por: Gu, Naibin, et al.
Publicado: (2025)
AT-MoE: Adaptive Task-planning Mixture of Experts via LoRA Approach
por: Li, Xurui, et al.
Publicado: (2024)
por: Li, Xurui, et al.
Publicado: (2024)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
por: Taniguchi, Rei, et al.
Publicado: (2026)
por: Taniguchi, Rei, et al.
Publicado: (2026)
Ejemplares similares
-
Gabliteration: Adaptive Multi-Directional Neural Weight Modification for Selective Behavioral Alteration in Large Language Models
por: Gülmez, Gökdeniz
Publicado: (2025) -
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
por: Jin, Peng, et al.
Publicado: (2024) -
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
por: Liu, Baihui, et al.
Publicado: (2026) -
LightMoE: Reducing Mixture-of-Experts Redundancy through Expert Replacing
por: Hao, Jiawei, et al.
Publicado: (2026) -
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
por: Wei, Jia, et al.
Publicado: (2026)