Towards Efficient Replay in Federated Incremental Learning

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Li, Yichen, Li, Qunwei, Wang, Haozhao, Li, Ruixuan, Zhong, Wenliang, Zhang, Guannan
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866914820256694272
author Li, Yichen
Li, Qunwei
Wang, Haozhao
Li, Ruixuan
Zhong, Wenliang
Zhang, Guannan
author_facet Li, Yichen
Li, Qunwei
Wang, Haozhao
Li, Ruixuan
Zhong, Wenliang
Zhang, Guannan
contents In Federated Learning (FL), the data in each client is typically assumed fixed or static. However, data often comes in an incremental manner in real-world applications, where the data domain may increase dynamically. In this work, we study catastrophic forgetting with data heterogeneity in Federated Incremental Learning (FIL) scenarios where edge clients may lack enough storage space to retain full data. We propose to employ a simple, generic framework for FIL named Re-Fed, which can coordinate each client to cache important samples for replay. More specifically, when a new task arrives, each client first caches selected previous samples based on their global and local importance. Then, the client trains the local model with both the cached samples and the samples from the new task. Theoretically, we analyze the ability of Re-Fed to discover important samples for replay thus alleviating the catastrophic forgetting problem. Moreover, we empirically show that Re-Fed achieves competitive performance compared to state-of-the-art methods.
format Preprint
id arxiv_https___arxiv_org_abs_2403_05890
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Towards Efficient Replay in Federated Incremental Learning
Li, Yichen
Li, Qunwei
Wang, Haozhao
Li, Ruixuan
Zhong, Wenliang
Zhang, Guannan
Machine Learning
Distributed, Parallel, and Cluster Computing
In Federated Learning (FL), the data in each client is typically assumed fixed or static. However, data often comes in an incremental manner in real-world applications, where the data domain may increase dynamically. In this work, we study catastrophic forgetting with data heterogeneity in Federated Incremental Learning (FIL) scenarios where edge clients may lack enough storage space to retain full data. We propose to employ a simple, generic framework for FIL named Re-Fed, which can coordinate each client to cache important samples for replay. More specifically, when a new task arrives, each client first caches selected previous samples based on their global and local importance. Then, the client trains the local model with both the cached samples and the samples from the new task. Theoretically, we analyze the ability of Re-Fed to discover important samples for replay thus alleviating the catastrophic forgetting problem. Moreover, we empirically show that Re-Fed achieves competitive performance compared to state-of-the-art methods.
title Towards Efficient Replay in Federated Incremental Learning
topic Machine Learning
Distributed, Parallel, and Cluster Computing
url https://arxiv.org/abs/2403.05890