Adaptive Model Ensemble for Continual Learning

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Mao, Yuchuan, Gao, Zhi, Fan, Xiaomeng, Wu, Yuwei, Jia, Yunde, Jing, Chenchen
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866914053851447296
author Mao, Yuchuan
Gao, Zhi
Fan, Xiaomeng
Wu, Yuwei
Jia, Yunde
Jing, Chenchen
author_facet Mao, Yuchuan
Gao, Zhi
Fan, Xiaomeng
Wu, Yuwei
Jia, Yunde
Jing, Chenchen
contents Model ensemble is an effective strategy in continual learning, which alleviates catastrophic forgetting by interpolating model parameters, achieving knowledge fusion learned from different tasks. However, existing model ensemble methods usually encounter the knowledge conflict issue at task and layer levels, causing compromised learning performance in both old and new tasks. To solve this issue, we propose meta-weight-ensembler that adaptively fuses knowledge of different tasks for continual learning. Concretely, we employ a mixing coefficient generator trained via meta-learning to generate appropriate mixing coefficients for model ensemble to address the task-level knowledge conflict. The mixing coefficient is individually generated for each layer to address the layer-level knowledge conflict. In this way, we learn the prior knowledge about adaptively accumulating knowledge of different tasks in a fused model, achieving efficient learning in both old and new tasks. Meta-weight-ensembler can be flexibly combined with existing continual learning methods to boost their ability of alleviating catastrophic forgetting. Experiments on multiple continual learning datasets show that meta-weight-ensembler effectively alleviates catastrophic forgetting and achieves state-of-the-art performance.
format Preprint
id arxiv_https___arxiv_org_abs_2509_19819
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Adaptive Model Ensemble for Continual Learning
Mao, Yuchuan
Gao, Zhi
Fan, Xiaomeng
Wu, Yuwei
Jia, Yunde
Jing, Chenchen
Computer Vision and Pattern Recognition
Model ensemble is an effective strategy in continual learning, which alleviates catastrophic forgetting by interpolating model parameters, achieving knowledge fusion learned from different tasks. However, existing model ensemble methods usually encounter the knowledge conflict issue at task and layer levels, causing compromised learning performance in both old and new tasks. To solve this issue, we propose meta-weight-ensembler that adaptively fuses knowledge of different tasks for continual learning. Concretely, we employ a mixing coefficient generator trained via meta-learning to generate appropriate mixing coefficients for model ensemble to address the task-level knowledge conflict. The mixing coefficient is individually generated for each layer to address the layer-level knowledge conflict. In this way, we learn the prior knowledge about adaptively accumulating knowledge of different tasks in a fused model, achieving efficient learning in both old and new tasks. Meta-weight-ensembler can be flexibly combined with existing continual learning methods to boost their ability of alleviating catastrophic forgetting. Experiments on multiple continual learning datasets show that meta-weight-ensembler effectively alleviates catastrophic forgetting and achieves state-of-the-art performance.
title Adaptive Model Ensemble for Continual Learning
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2509.19819