Guardado en:
Detalles Bibliográficos
Autores principales: Michel, Nicolas, Wang, Maorong, He, Jiangpeng, Yamasaki, Toshihiko
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:https://arxiv.org/abs/2502.18762
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866913882301267968
author Michel, Nicolas
Wang, Maorong
He, Jiangpeng
Yamasaki, Toshihiko
author_facet Michel, Nicolas
Wang, Maorong
He, Jiangpeng
Yamasaki, Toshihiko
contents Continual Learning (CL) aims to learn from a non-stationary data stream where the underlying distribution changes over time. While recent advances have produced efficient memory-free methods in the offline CL (offCL) setting, where tasks are known in advance and data can be revisited, online CL (onCL) remains dominated by memory-based approaches. The transition from offCL to onCL is challenging, as many offline methods rely on (1) prior knowledge of task boundaries and (2) sophisticated scheduling or optimization schemes, both of which are unavailable when data arrives sequentially and can be seen only once. In this paper, we investigate the adaptation of state-of-the-art memory-free offCL methods to the online setting. We first show that augmenting these methods with lightweight prototypes significantly improves performance, albeit at the cost of increased Gradient Imbalance, resulting in a biased learning towards earlier tasks. To address this issue, we introduce Fine-Grained Hypergradients, an online mechanism for rebalancing gradient updates during training. Our experiments demonstrate that the synergy between prototype memory and hypergradient reweighting substantially enhances the performance of memory-free methods in onCL and surpasses onCL baselines. Code will be released upon acceptance.
format Preprint
id arxiv_https___arxiv_org_abs_2502_18762
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle From Offline to Online Memory-Free and Task-Free Continual Learning via Fine-Grained Hypergradients
Michel, Nicolas
Wang, Maorong
He, Jiangpeng
Yamasaki, Toshihiko
Machine Learning
Artificial Intelligence
Continual Learning (CL) aims to learn from a non-stationary data stream where the underlying distribution changes over time. While recent advances have produced efficient memory-free methods in the offline CL (offCL) setting, where tasks are known in advance and data can be revisited, online CL (onCL) remains dominated by memory-based approaches. The transition from offCL to onCL is challenging, as many offline methods rely on (1) prior knowledge of task boundaries and (2) sophisticated scheduling or optimization schemes, both of which are unavailable when data arrives sequentially and can be seen only once. In this paper, we investigate the adaptation of state-of-the-art memory-free offCL methods to the online setting. We first show that augmenting these methods with lightweight prototypes significantly improves performance, albeit at the cost of increased Gradient Imbalance, resulting in a biased learning towards earlier tasks. To address this issue, we introduce Fine-Grained Hypergradients, an online mechanism for rebalancing gradient updates during training. Our experiments demonstrate that the synergy between prototype memory and hypergradient reweighting substantially enhances the performance of memory-free methods in onCL and surpasses onCL baselines. Code will be released upon acceptance.
title From Offline to Online Memory-Free and Task-Free Continual Learning via Fine-Grained Hypergradients
topic Machine Learning
Artificial Intelligence
url https://arxiv.org/abs/2502.18762