Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution
Fuente:
arXiv
Saved in:
| Main Authors: | Qiu, Haiyun, Wu, Xingyu, Tan, Kay Chen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fine-Grained Model Merging via Modular Expert Recombination
by: Qiu, Haiyun, et al.
Published: (2026)
by: Qiu, Haiyun, et al.
Published: (2026)
HM3: Hierarchical Multi-Objective Model Merging for Pretrained Models
by: Zhou, Yu, et al.
Published: (2024)
by: Zhou, Yu, et al.
Published: (2024)
Diversity-Aware Policy Optimization for Large Language Model Reasoning
by: Yao, Jian, et al.
Published: (2025)
by: Yao, Jian, et al.
Published: (2025)
Expert Merging: Model Merging with Unsupervised Expert Alignment and Importance-Guided Layer Chunking
by: Zhang, Dengming, et al.
Published: (2025)
by: Zhang, Dengming, et al.
Published: (2025)
Soft Merging of Experts with Adaptive Routing
by: Muqeeth, Mohammed, et al.
Published: (2023)
by: Muqeeth, Mohammed, et al.
Published: (2023)
Large Language Model-Enhanced Algorithm Selection: Towards Comprehensive Algorithm Representation
by: Wu, Xingyu, et al.
Published: (2023)
by: Wu, Xingyu, et al.
Published: (2023)
Sparsity-Aware Evolution for Model Merging
by: Zhang, Huan, et al.
Published: (2026)
by: Zhang, Huan, et al.
Published: (2026)
MINGLE: Mixture of Null-Space Gated Low-Rank Experts for Test-Time Continual Model Merging
by: Qiu, Zihuan, et al.
Published: (2025)
by: Qiu, Zihuan, et al.
Published: (2025)
Design Principle Transfer in Neural Architecture Search via Large Language Models
by: Zhou, Xun, et al.
Published: (2024)
by: Zhou, Xun, et al.
Published: (2024)
Expert Merging in Sparse Mixture of Experts with Nash Bargaining
by: Nguyen, Dung V., et al.
Published: (2025)
by: Nguyen, Dung V., et al.
Published: (2025)
MergeMoE: Efficient Compression of MoE Models via Expert Output Merging
by: Miao, Ruijie, et al.
Published: (2025)
by: Miao, Ruijie, et al.
Published: (2025)
CAMEx: Curvature-aware Merging of Experts
by: Nguyen, Dung V., et al.
Published: (2025)
by: Nguyen, Dung V., et al.
Published: (2025)
LLM Cannot Discover Causality, and Should Be Restricted to Non-Decisional Support in Causal Discovery
by: Wu, Xingyu, et al.
Published: (2025)
by: Wu, Xingyu, et al.
Published: (2025)
Pareto Merging: Multi-Objective Optimization for Preference-Aware Model Merging
by: Chen, Weiyu, et al.
Published: (2024)
by: Chen, Weiyu, et al.
Published: (2024)
FAME: Forecasting Academic Impact via Continuous-Time Manifold Evolution
by: Ding, Jianrong, et al.
Published: (2026)
by: Ding, Jianrong, et al.
Published: (2026)
Vanishing Feature: Diagnosing Model Merging and Beyond
by: Qu, Xingyu, et al.
Published: (2024)
by: Qu, Xingyu, et al.
Published: (2024)
Toward a Holistic Approach to Continual Model Merging
by: Phan, Hoang, et al.
Published: (2025)
by: Phan, Hoang, et al.
Published: (2025)
Learning More Generalized Experts by Merging Experts in Mixture-of-Experts
by: Park, Sejik
Published: (2024)
by: Park, Sejik
Published: (2024)
Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning
by: Hiroshima, Kei, et al.
Published: (2026)
by: Hiroshima, Kei, et al.
Published: (2026)
FedMerge: Federated Personalization via Model Merging
by: Chen, Shutong, et al.
Published: (2025)
by: Chen, Shutong, et al.
Published: (2025)
Merge before Forget: A Single LoRA Continual Learning via Continual Merging
by: Qiao, Fuli, et al.
Published: (2025)
by: Qiao, Fuli, et al.
Published: (2025)
HyperMoE: Towards Better Mixture of Experts via Transferring Among Experts
by: Zhao, Hao, et al.
Published: (2024)
by: Zhao, Hao, et al.
Published: (2024)
PASs-MoE: Mitigating Misaligned Co-drift among Router and Experts via Pathway Activation Subspaces for Continual Learning
by: Hou, Zhiyan, et al.
Published: (2026)
by: Hou, Zhiyan, et al.
Published: (2026)
Adaptive Substructure-Aware Expert Model for Molecular Property Prediction
by: Jiang, Tianyi, et al.
Published: (2025)
by: Jiang, Tianyi, et al.
Published: (2025)
Faster, Smaller, and Smarter: Task-Aware Expert Merging for Online MoE Inference
by: Han, Ziyi, et al.
Published: (2025)
by: Han, Ziyi, et al.
Published: (2025)
Mitigating the Backdoor Effect for Multi-Task Model Merging via Safety-Aware Subspace
by: Yang, Jinluan, et al.
Published: (2024)
by: Yang, Jinluan, et al.
Published: (2024)
Sub-MoE: Efficient Mixture-of-Expert LLMs Compression via Subspace Expert Merging
by: Li, Lujun, et al.
Published: (2025)
by: Li, Lujun, et al.
Published: (2025)
Learning from B Cell Evolution: Adaptive Multi-Expert Diffusion for Antibody Design via Online Optimization
by: Feng, Hanqi, et al.
Published: (2025)
by: Feng, Hanqi, et al.
Published: (2025)
BD-Merging: Bias-Aware Dynamic Model Merging with Evidence-Guided Contrastive Learning
by: Xie, Yuhan, et al.
Published: (2026)
by: Xie, Yuhan, et al.
Published: (2026)
Null-Space Filtering for Data-Free Continual Model Merging: Preserving Stability, Promoting Plasticity
by: Qiu, Zihuan, et al.
Published: (2025)
by: Qiu, Zihuan, et al.
Published: (2025)
Data Augmentation for Continual RL via Adversarial Gradient Episodic Memory
by: Wu, Sihao, et al.
Published: (2024)
by: Wu, Sihao, et al.
Published: (2024)
Channel Merging: Preserving Specialization for Merged Experts
by: Zhang, Mingyang, et al.
Published: (2024)
by: Zhang, Mingyang, et al.
Published: (2024)
Merging Multi-Task Models via Weight-Ensembling Mixture of Experts
by: Tang, Anke, et al.
Published: (2024)
by: Tang, Anke, et al.
Published: (2024)
Merging Models on the Fly Without Retraining: A Sequential Approach to Scalable Continual Model Merging
by: Tang, Anke, et al.
Published: (2025)
by: Tang, Anke, et al.
Published: (2025)
CausalBench: A Comprehensive Benchmark for Causal Learning Capability of LLMs
by: Zhou, Yu, et al.
Published: (2024)
by: Zhou, Yu, et al.
Published: (2024)
Efficient Estimation for Longitudinal Networks via Adaptive Merging
by: Zhang, Haoran, et al.
Published: (2022)
by: Zhang, Haoran, et al.
Published: (2022)
Exploiting Task Relationships in Continual Learning via Transferability-Aware Task Embeddings
by: Wu, Yanru, et al.
Published: (2025)
by: Wu, Yanru, et al.
Published: (2025)
Unlock the Power of Algorithm Features: A Generalization Analysis for Algorithm Selection
by: Wu, Xingyu, et al.
Published: (2024)
by: Wu, Xingyu, et al.
Published: (2024)
Task-Aware Mixture-of-Experts for Time Series Analysis
by: Wu, Xingjian, et al.
Published: (2025)
by: Wu, Xingjian, et al.
Published: (2025)
Local Mixtures of Experts: Essentially Free Test-Time Training via Model Merging
by: Bertolissi, Ryo, et al.
Published: (2025)
by: Bertolissi, Ryo, et al.
Published: (2025)
Similar Items
-
Fine-Grained Model Merging via Modular Expert Recombination
by: Qiu, Haiyun, et al.
Published: (2026) -
HM3: Hierarchical Multi-Objective Model Merging for Pretrained Models
by: Zhou, Yu, et al.
Published: (2024) -
Diversity-Aware Policy Optimization for Large Language Model Reasoning
by: Yao, Jian, et al.
Published: (2025) -
Expert Merging: Model Merging with Unsupervised Expert Alignment and Importance-Guided Layer Chunking
by: Zhang, Dengming, et al.
Published: (2025) -
Soft Merging of Experts with Adaptive Routing
by: Muqeeth, Mohammed, et al.
Published: (2023)