Similarity-Aware Mixture-of-Experts for Data-Efficient Continual Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Mclaughlin, Connor, Lee, Nigel, Su, Lili |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalized Federated Learning via Feature Distribution Adaptation
by: Mclaughlin, Connor J., et al.
Published: (2024)
by: Mclaughlin, Connor J., et al.
Published: (2024)
On the Power of Source Screening for Learning Shared Feature Extractors
by: Wang, Leo Muxing, et al.
Published: (2026)
by: Wang, Leo Muxing, et al.
Published: (2026)
Fast and Robust State Estimation and Tracking via Hierarchical Learning
by: Mclaughlin, Connor, et al.
Published: (2023)
by: Mclaughlin, Connor, et al.
Published: (2023)
Continual Traffic Forecasting via Mixture of Experts
by: Lee, Sanghyun, et al.
Published: (2024)
by: Lee, Sanghyun, et al.
Published: (2024)
Load Balancing Mixture of Experts with Similarity Preserving Routers
by: Omi, Nabil, et al.
Published: (2025)
by: Omi, Nabil, et al.
Published: (2025)
Theory on Mixture-of-Experts in Continual Learning
by: Li, Hongbo, et al.
Published: (2024)
by: Li, Hongbo, et al.
Published: (2024)
Mixture of Experts Meets Prompt-Based Continual Learning
by: Le, Minh, et al.
Published: (2024)
by: Le, Minh, et al.
Published: (2024)
Klotski: Efficient Mixture-of-Expert Inference via Expert-Aware Multi-Batch Pipeline
by: Fang, Zhiyuan, et al.
Published: (2025)
by: Fang, Zhiyuan, et al.
Published: (2025)
PWC-MoE: Privacy-Aware Wireless Collaborative Mixture of Experts
by: Su, Yang, et al.
Published: (2025)
by: Su, Yang, et al.
Published: (2025)
Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning
by: Siddika, Fatema, et al.
Published: (2026)
by: Siddika, Fatema, et al.
Published: (2026)
FLAME: Adaptive Mixture-of-Experts for Continual Multimodal Multi-Task Learning
by: Han, Xing, et al.
Published: (2026)
by: Han, Xing, et al.
Published: (2026)
Learning More Generalized Experts by Merging Experts in Mixture-of-Experts
by: Park, Sejik
Published: (2024)
by: Park, Sejik
Published: (2024)
CRAFT: Fine-Grained Cost-Aware Expert Replication For Efficient Mixture-of-Experts Serving
by: Zhao, Adrian, et al.
Published: (2026)
by: Zhao, Adrian, et al.
Published: (2026)
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping
by: Huang, Ziye, et al.
Published: (2024)
by: Huang, Ziye, et al.
Published: (2024)
Mixtures of SubExperts for Large Language Continual Learning
by: Kang, Haeyong
Published: (2025)
by: Kang, Haeyong
Published: (2025)
Task-Aware Mixture-of-Experts for Time Series Analysis
by: Wu, Xingjian, et al.
Published: (2025)
by: Wu, Xingjian, et al.
Published: (2025)
Efficiently Editing Mixture-of-Experts Models with Compressed Experts
by: He, Yifei, et al.
Published: (2025)
by: He, Yifei, et al.
Published: (2025)
Mixture of Predefined Experts: Maximizing Data Usage on Vertical Federated Learning
by: Irureta, Jon, et al.
Published: (2026)
by: Irureta, Jon, et al.
Published: (2026)
Topology-Aware Multiscale Mixture of Experts for Efficient Molecular Property Prediction
by: Nguyen, Long D., et al.
Published: (2026)
by: Nguyen, Long D., et al.
Published: (2026)
Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference
by: Liu, Baihui, et al.
Published: (2026)
by: Liu, Baihui, et al.
Published: (2026)
Shift Happens: Mixture of Experts based Continual Adaptation in Federated Learning
by: Bhope, Rahul Atul, et al.
Published: (2025)
by: Bhope, Rahul Atul, et al.
Published: (2025)
Expert Routing with Synthetic Data for Continual Learning
by: Byun, Yewon, et al.
Published: (2024)
by: Byun, Yewon, et al.
Published: (2024)
Scene-Adaptive Continual Learning for CSI-based Human Activity Recognition with Mixture of Experts
by: Zheng, Wenhan, et al.
Published: (2026)
by: Zheng, Wenhan, et al.
Published: (2026)
One-Prompt Strikes Back: Sparse Mixture of Experts for Prompt-based Continual Learning
by: Le, Minh, et al.
Published: (2025)
by: Le, Minh, et al.
Published: (2025)
Efficient Diffusion Transformer Policies with Mixture of Expert Denoisers for Multitask Learning
by: Reuss, Moritz, et al.
Published: (2024)
by: Reuss, Moritz, et al.
Published: (2024)
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
by: Lin, Xi Victoria, et al.
Published: (2024)
by: Lin, Xi Victoria, et al.
Published: (2024)
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
by: Belenki, Lior, et al.
Published: (2025)
by: Belenki, Lior, et al.
Published: (2025)
Expert Upcycling: Shifting the Compute-Efficient Frontier of Mixture-of-Experts
by: Dwivedi, Chaitanya, et al.
Published: (2026)
by: Dwivedi, Chaitanya, et al.
Published: (2026)
Mixture-of-Clustered-Experts: Advancing Expert Specialization and Generalization in Instruction Tuning
by: Eo, Sugyeong, et al.
Published: (2025)
by: Eo, Sugyeong, et al.
Published: (2025)
Is Temperature Sample Efficient for Softmax Gaussian Mixture of Experts?
by: Nguyen, Huy, et al.
Published: (2024)
by: Nguyen, Huy, et al.
Published: (2024)
Variational Inference, Entropy, and Orthogonality: A Unified Theory of Mixture-of-Experts
by: Su, Ye, et al.
Published: (2026)
by: Su, Ye, et al.
Published: (2026)
From Molecules to Mixtures: Learning Representations of Olfactory Mixture Similarity using Inductive Biases
by: Tom, Gary, et al.
Published: (2025)
by: Tom, Gary, et al.
Published: (2025)
Separation and Collaboration: Two-Level Routing Grouped Mixture-of-Experts for Multi-Domain Continual Learning
by: Zhou, Jialu, et al.
Published: (2025)
by: Zhou, Jialu, et al.
Published: (2025)
Little by Little: Continual Learning via Incremental Mixture of Rank-1 Associative Memory Experts
by: Lu, Haodong, et al.
Published: (2025)
by: Lu, Haodong, et al.
Published: (2025)
Capacity-Aware Mixture Law Enables Efficient LLM Data Optimization
by: Li, Jingwei, et al.
Published: (2026)
by: Li, Jingwei, et al.
Published: (2026)
H+: An Efficient Similarity-Aware Aggregation for Byzantine Resilient Federated Learning
by: Zuo, Shiyuan, et al.
Published: (2025)
by: Zuo, Shiyuan, et al.
Published: (2025)
Mixture-of-Experts for Distributed Edge Computing with Channel-Aware Gating Function
by: Song, Qiuchen, et al.
Published: (2025)
by: Song, Qiuchen, et al.
Published: (2025)
Dynamic Mixture-of-Experts for Incremental Graph Learning
by: Kong, Lecheng, et al.
Published: (2025)
by: Kong, Lecheng, et al.
Published: (2025)
Anchor-MoE: A Mean-Anchored Mixture of Experts For Probabilistic Regression
by: Su, Baozhuo, et al.
Published: (2025)
by: Su, Baozhuo, et al.
Published: (2025)
Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs
by: Benazir, Afsara, et al.
Published: (2026)
by: Benazir, Afsara, et al.
Published: (2026)
Similar Items
-
Personalized Federated Learning via Feature Distribution Adaptation
by: Mclaughlin, Connor J., et al.
Published: (2024) -
On the Power of Source Screening for Learning Shared Feature Extractors
by: Wang, Leo Muxing, et al.
Published: (2026) -
Fast and Robust State Estimation and Tracking via Hierarchical Learning
by: Mclaughlin, Connor, et al.
Published: (2023) -
Continual Traffic Forecasting via Mixture of Experts
by: Lee, Sanghyun, et al.
Published: (2024) -
Load Balancing Mixture of Experts with Similarity Preserving Routers
by: Omi, Nabil, et al.
Published: (2025)