MoTE: Mixture of Task-specific Experts for Pre-Trained ModelBased Class-incremental Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Linjie, Wu, Zhenyu, Ji, Yang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models
von: Romero, Miguel, et al.
Veröffentlicht: (2025)
von: Romero, Miguel, et al.
Veröffentlicht: (2025)
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
Quantum-Gated Task-interaction Knowledge Distillation for Pre-trained Model-based Class-Incremental Learning
von: Li, Linjie, et al.
Veröffentlicht: (2026)
von: Li, Linjie, et al.
Veröffentlicht: (2026)
Mixture of insighTful Experts (MoTE): The Synergy of Thought Chains and Expert Mixtures in Self-Alignment
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
von: Liu, Zhili, et al.
Veröffentlicht: (2024)
Mixture of Noise for Pre-Trained Model-Based Class-Incremental Learning
von: Jiang, Kai, et al.
Veröffentlicht: (2025)
von: Jiang, Kai, et al.
Veröffentlicht: (2025)
MoRE: A Mixture of Low-Rank Experts for Adaptive Multi-Task Learning
von: Zhang, Dacao, et al.
Veröffentlicht: (2025)
von: Zhang, Dacao, et al.
Veröffentlicht: (2025)
INCPrompt: Task-Aware incremental Prompting for Rehearsal-Free Class-incremental Learning
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2024)
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
PreMoE: Proactive Inference for Efficient Mixture-of-Experts
von: Pei, Zehua, et al.
Veröffentlicht: (2025)
von: Pei, Zehua, et al.
Veröffentlicht: (2025)
DA-MoE: Towards Dynamic Expert Allocation for Mixture-of-Experts Models
von: Aghdam, Maryam Akhavan, et al.
Veröffentlicht: (2024)
von: Aghdam, Maryam Akhavan, et al.
Veröffentlicht: (2024)
Chain-of-Experts: Unlocking the Communication Power of Mixture-of-Experts Models
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
von: Wang, Zihan, et al.
Veröffentlicht: (2025)
L-MoE: End-to-End Training of a Lightweight Mixture of Low-Rank Adaptation Experts
von: Ji, Shihao, et al.
Veröffentlicht: (2025)
von: Ji, Shihao, et al.
Veröffentlicht: (2025)
Task-Aware Mixture-of-Experts for Time Series Analysis
von: Wu, Xingjian, et al.
Veröffentlicht: (2025)
von: Wu, Xingjian, et al.
Veröffentlicht: (2025)
eMoE: Task-aware Memory Efficient Mixture-of-Experts-Based (MoE) Model Inference
von: Tairin, Suraiya, et al.
Veröffentlicht: (2025)
von: Tairin, Suraiya, et al.
Veröffentlicht: (2025)
PTMs-TSCIL Pre-Trained Models Based Class-Incremental Learning
von: Wu, Yuanlong, et al.
Veröffentlicht: (2025)
von: Wu, Yuanlong, et al.
Veröffentlicht: (2025)
HeterMoE: Efficient Training of Mixture-of-Experts Models on Heterogeneous GPUs
von: Wu, Yongji, et al.
Veröffentlicht: (2025)
von: Wu, Yongji, et al.
Veröffentlicht: (2025)
Mixture-of-Experts Operator Transformer for Large-Scale PDE Pre-Training
von: Wang, Hong, et al.
Veröffentlicht: (2025)
von: Wang, Hong, et al.
Veröffentlicht: (2025)
MoEMeta: Mixture-of-Experts Meta Learning for Few-Shot Relational Learning
von: Wu, Han, et al.
Veröffentlicht: (2025)
von: Wu, Han, et al.
Veröffentlicht: (2025)
Adaptive Shared Experts with LoRA-Based Mixture of Experts for Multi-Task Learning
von: Yang, Minghao, et al.
Veröffentlicht: (2025)
von: Yang, Minghao, et al.
Veröffentlicht: (2025)
Symphony-MoE: Harmonizing Disparate Pre-trained Models into a Coherent Mixture-of-Experts
von: Wang, Qi, et al.
Veröffentlicht: (2025)
von: Wang, Qi, et al.
Veröffentlicht: (2025)
MoBE: Mixture-of-Basis-Experts for Compressing MoE-based LLMs
von: Chen, Xiaodong, et al.
Veröffentlicht: (2025)
von: Chen, Xiaodong, et al.
Veröffentlicht: (2025)
LAER-MoE: Load-Adaptive Expert Re-layout for Efficient Mixture-of-Experts Training
von: Liu, Xinyi, et al.
Veröffentlicht: (2026)
von: Liu, Xinyi, et al.
Veröffentlicht: (2026)
MoE-DisCo:Low Economy Cost Training Mixture-of-Experts Models
von: Ye, Xin, et al.
Veröffentlicht: (2026)
von: Ye, Xin, et al.
Veröffentlicht: (2026)
Input Domain Aware MoE: Decoupling Routing Decisions from Task Optimization in Mixture of Experts
von: Hua, Yongxiang, et al.
Veröffentlicht: (2025)
von: Hua, Yongxiang, et al.
Veröffentlicht: (2025)
Dual Prototypes for Adaptive Pre-Trained Model in Class-Incremental Learning
von: Xu, Zhiming, et al.
Veröffentlicht: (2024)
von: Xu, Zhiming, et al.
Veröffentlicht: (2024)
DirMoE: Dirichlet-routed Mixture of Experts
von: Vahidi, Amirhossein, et al.
Veröffentlicht: (2026)
von: Vahidi, Amirhossein, et al.
Veröffentlicht: (2026)
Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2023)
von: Hendawy, Ahmed, et al.
Veröffentlicht: (2023)
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
von: Lin, Xi Victoria, et al.
Veröffentlicht: (2024)
von: Lin, Xi Victoria, et al.
Veröffentlicht: (2024)
Horseshoe Mixtures-of-Experts (HS-MoE)
von: Polson, Nick, et al.
Veröffentlicht: (2026)
von: Polson, Nick, et al.
Veröffentlicht: (2026)
MoDEx: Mixture of Depth-specific Experts for Multivariate Long-term Time Series Forecasting
von: Yoon, Hyekyung, et al.
Veröffentlicht: (2026)
von: Yoon, Hyekyung, et al.
Veröffentlicht: (2026)
MoBiE: Efficient Inference of Mixture of Binary Experts under Post-Training Quantization
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)
von: Zhao, Zhixiong, et al.
Veröffentlicht: (2026)
MoNTA: Accelerating Mixture-of-Experts Training with Network-Traffc-Aware Parallel Optimization
von: Guo, Jingming, et al.
Veröffentlicht: (2024)
von: Guo, Jingming, et al.
Veröffentlicht: (2024)
Elastic MoE: Unlocking the Inference-Time Scalability of Mixture-of-Experts
von: Gu, Naibin, et al.
Veröffentlicht: (2025)
von: Gu, Naibin, et al.
Veröffentlicht: (2025)
QuantMoE-Bench: Examining Post-Training Quantization for Mixture-of-Experts
von: Li, Pingzhi, et al.
Veröffentlicht: (2024)
von: Li, Pingzhi, et al.
Veröffentlicht: (2024)
MoLAE: Mixture of Latent Experts for Parameter-Efficient Language Models
von: Liu, Zehua, et al.
Veröffentlicht: (2025)
von: Liu, Zehua, et al.
Veröffentlicht: (2025)
MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
von: Jin, Peng, et al.
Veröffentlicht: (2024)
von: Jin, Peng, et al.
Veröffentlicht: (2024)
BigMac: A Communication-Efficient Mixture-of-Experts Model Structure for Fast Training and Inference
von: Jin, Zewen, et al.
Veröffentlicht: (2025)
von: Jin, Zewen, et al.
Veröffentlicht: (2025)
Each Rank Could be an Expert: Single-Ranked Mixture of Experts LoRA for Multi-Task Learning
von: Zhao, Ziyu, et al.
Veröffentlicht: (2025)
von: Zhao, Ziyu, et al.
Veröffentlicht: (2025)
MoC-System: Efficient Fault Tolerance for Sparse Mixture-of-Experts Model Training
von: Cai, Weilin, et al.
Veröffentlicht: (2024)
von: Cai, Weilin, et al.
Veröffentlicht: (2024)
MoDE: A Mixture-of-Experts Model with Mutual Distillation among the Experts
von: Xie, Zhitian, et al.
Veröffentlicht: (2024)
von: Xie, Zhitian, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Beyond instruction-conditioning, MoTE: Mixture of Task Experts for Multi-task Embedding Models
von: Romero, Miguel, et al.
Veröffentlicht: (2025) -
MoTE: Mixture of Ternary Experts for Memory-efficient Large Multimodal Models
von: Wang, Hongyu, et al.
Veröffentlicht: (2025) -
Quantum-Gated Task-interaction Knowledge Distillation for Pre-trained Model-based Class-Incremental Learning
von: Li, Linjie, et al.
Veröffentlicht: (2026) -
Mixture of insighTful Experts (MoTE): The Synergy of Thought Chains and Expert Mixtures in Self-Alignment
von: Liu, Zhili, et al.
Veröffentlicht: (2024) -
Mixture of Noise for Pre-Trained Model-Based Class-Incremental Learning
von: Jiang, Kai, et al.
Veröffentlicht: (2025)