PT-MoE: An Efficient Finetuning Framework for Integrating Mixture-of-Experts into Prompt Tuning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Zongqian, Su, Yixuan, Collier, Nigel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Survey on Prompt Tuning
por: Li, Zongqian, et al.
Publicado: (2025)
por: Li, Zongqian, et al.
Publicado: (2025)
500xCompressor: Generalized Prompt Compression for Large Language Models
por: Li, Zongqian, et al.
Publicado: (2024)
por: Li, Zongqian, et al.
Publicado: (2024)
Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks
por: Li, Zongqian, et al.
Publicado: (2026)
por: Li, Zongqian, et al.
Publicado: (2026)
Prompt Compression for Large Language Models: A Survey
por: Li, Zongqian, et al.
Publicado: (2024)
por: Li, Zongqian, et al.
Publicado: (2024)
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
por: Tang, Yehui, et al.
Publicado: (2025)
por: Tang, Yehui, et al.
Publicado: (2025)
ReasonGraph: Visualisation of Reasoning Paths
por: Li, Zongqian, et al.
Publicado: (2025)
por: Li, Zongqian, et al.
Publicado: (2025)
MH-MoE: Multi-Head Mixture-of-Experts
por: Huang, Shaohan, et al.
Publicado: (2024)
por: Huang, Shaohan, et al.
Publicado: (2024)
$\infty$-MoE: Generalizing Mixture of Experts to Infinite Experts
por: Takashiro, Shota, et al.
Publicado: (2026)
por: Takashiro, Shota, et al.
Publicado: (2026)
Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts
por: Wu, Haoyuan, et al.
Publicado: (2025)
por: Wu, Haoyuan, et al.
Publicado: (2025)
EPS-MoE: Expert Pipeline Scheduler for Cost-Efficient MoE Inference
por: Qian, Yulei, et al.
Publicado: (2024)
por: Qian, Yulei, et al.
Publicado: (2024)
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
por: Liu, Baihui, et al.
Publicado: (2026)
por: Liu, Baihui, et al.
Publicado: (2026)
SEER-MoE: Sparse Expert Efficiency through Regularization for Mixture-of-Experts
por: Muzio, Alexandre, et al.
Publicado: (2024)
por: Muzio, Alexandre, et al.
Publicado: (2024)
Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient
por: Ludziejewski, Jan, et al.
Publicado: (2025)
por: Ludziejewski, Jan, et al.
Publicado: (2025)
MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts
por: Pióro, Maciej, et al.
Publicado: (2024)
por: Pióro, Maciej, et al.
Publicado: (2024)
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling
por: Jiang, Fan, et al.
Publicado: (2026)
por: Jiang, Fan, et al.
Publicado: (2026)
S2MoE: Robust Sparse Mixture of Experts via Stochastic Learning
por: Do, Giang, et al.
Publicado: (2025)
por: Do, Giang, et al.
Publicado: (2025)
Uni-MoE: Scaling Unified Multimodal LLMs with Mixture of Experts
por: Li, Yunxin, et al.
Publicado: (2024)
por: Li, Yunxin, et al.
Publicado: (2024)
MoE-Sieve: Routing-Guided LoRA for Efficient MoE Fine-Tuning
por: Manzoni, Andrea
Publicado: (2026)
por: Manzoni, Andrea
Publicado: (2026)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
por: Gu, Naibin, et al.
Publicado: (2025)
por: Gu, Naibin, et al.
Publicado: (2025)
OmniMoE: An Efficient MoE by Orchestrating Atomic Experts at Scale
por: Shi, Jingze, et al.
Publicado: (2026)
por: Shi, Jingze, et al.
Publicado: (2026)
COFFEE: A Contrastive Oracle-Free Framework for Event Extraction
por: Zhang, Meiru, et al.
Publicado: (2023)
por: Zhang, Meiru, et al.
Publicado: (2023)
Linear-MoE: Linear Sequence Modeling Meets Mixture-of-Experts
por: Sun, Weigao, et al.
Publicado: (2025)
por: Sun, Weigao, et al.
Publicado: (2025)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
por: Liu, Yang, et al.
Publicado: (2026)
por: Liu, Yang, et al.
Publicado: (2026)
Breaking Training Bottlenecks: Effective and Stable Reinforcement Learning for Coding Models
por: Li, Zongqian, et al.
Publicado: (2026)
por: Li, Zongqian, et al.
Publicado: (2026)
Med-MoE: Mixture of Domain-Specific Experts for Lightweight Medical Vision-Language Models
por: Jiang, Songtao, et al.
Publicado: (2024)
por: Jiang, Songtao, et al.
Publicado: (2024)
Unlocking Structure Measuring: Introducing PDD, an Automatic Metric for Positional Discourse Coherence
por: Liu, Yinhong, et al.
Publicado: (2024)
por: Liu, Yinhong, et al.
Publicado: (2024)
$μ$-MoE: Test-Time Pruning as Micro-Grained Mixture-of-Experts
por: Koike-Akino, Toshiaki, et al.
Publicado: (2025)
por: Koike-Akino, Toshiaki, et al.
Publicado: (2025)
MoLEx: Mixture of Layer Experts for Finetuning with Sparse Upcycling
por: Teo, Rachel S. Y., et al.
Publicado: (2025)
por: Teo, Rachel S. Y., et al.
Publicado: (2025)
Skywork-MoE: A Deep Dive into Training Techniques for Mixture-of-Experts Language Models
por: Wei, Tianwen, et al.
Publicado: (2024)
por: Wei, Tianwen, et al.
Publicado: (2024)
DIVE into MoE: Diversity-Enhanced Reconstruction of Large Language Models from Dense into Mixture-of-Experts
por: Feng, Yuchen, et al.
Publicado: (2025)
por: Feng, Yuchen, et al.
Publicado: (2025)
MoE-DiffuSeq: Enhancing Long-Document Diffusion Models with Sparse Attention and Mixture of Experts
por: Christoforos, Alexandros, et al.
Publicado: (2025)
por: Christoforos, Alexandros, et al.
Publicado: (2025)
LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training
por: Zhu, Tong, et al.
Publicado: (2024)
por: Zhu, Tong, et al.
Publicado: (2024)
X-MoE: Enabling Scalable Training for Emerging Mixture-of-Experts Architectures on HPC Platforms
por: Yuan, Yueming, et al.
Publicado: (2025)
por: Yuan, Yueming, et al.
Publicado: (2025)
$\texttt{MoE-RBench}$: Towards Building Reliable Language Models with Sparse Mixture-of-Experts
por: Chen, Guanjie, et al.
Publicado: (2024)
por: Chen, Guanjie, et al.
Publicado: (2024)
MoE-LPR: Multilingual Extension of Large Language Models through Mixture-of-Experts with Language Priors Routing
por: Zhou, Hao, et al.
Publicado: (2024)
por: Zhou, Hao, et al.
Publicado: (2024)
MoE-Prism: Disentangling Monolithic Experts for Elastic MoE Services via Model-System Co-Designs
por: Xia, Xinfeng, et al.
Publicado: (2025)
por: Xia, Xinfeng, et al.
Publicado: (2025)
FLAME-MoE: A Transparent End-to-End Research Platform for Mixture-of-Experts Language Models
por: Kang, Hao, et al.
Publicado: (2025)
por: Kang, Hao, et al.
Publicado: (2025)
Attention Instruction: Amplifying Attention in the Middle via Prompting
por: Zhang, Meiru, et al.
Publicado: (2024)
por: Zhang, Meiru, et al.
Publicado: (2024)
BLR-MoE: Boosted Language-Routing Mixture of Experts for Domain-Robust Multilingual E2E ASR
por: Ma, Guodong, et al.
Publicado: (2025)
por: Ma, Guodong, et al.
Publicado: (2025)
MoMoE: Mixture of Moderation Experts Framework for AI-Assisted Online Governance
por: Goyal, Agam, et al.
Publicado: (2025)
por: Goyal, Agam, et al.
Publicado: (2025)
Ejemplares similares
-
A Survey on Prompt Tuning
por: Li, Zongqian, et al.
Publicado: (2025) -
500xCompressor: Generalized Prompt Compression for Large Language Models
por: Li, Zongqian, et al.
Publicado: (2024) -
Flexi-LoRA with Input-Adaptive Ranks: Efficient Finetuning for Speech and Reasoning Tasks
por: Li, Zongqian, et al.
Publicado: (2026) -
Prompt Compression for Large Language Models: A Survey
por: Li, Zongqian, et al.
Publicado: (2024) -
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity
por: Tang, Yehui, et al.
Publicado: (2025)