Machine Unlearning for Streaming Forgetting

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Shen, Shaofei, Zhang, Chenhao, Zhao, Yawen, Bialkowski, Alina, Chen, Weitong, Xu, Miao
Formato: Preprint
Publicado: 2025
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866912493504299008
author Shen, Shaofei
Zhang, Chenhao
Zhao, Yawen
Bialkowski, Alina
Chen, Weitong
Xu, Miao
author_facet Shen, Shaofei
Zhang, Chenhao
Zhao, Yawen
Bialkowski, Alina
Chen, Weitong
Xu, Miao
contents Machine unlearning aims to remove knowledge of the specific training data in a well-trained model. Currently, machine unlearning methods typically handle all forgetting data in a single batch, removing the corresponding knowledge all at once upon request. However, in practical scenarios, requests for data removal often arise in a streaming manner rather than in a single batch, leading to reduced efficiency and effectiveness in existing methods. Such challenges of streaming forgetting have not been the focus of much research. In this paper, to address the challenges of performance maintenance, efficiency, and data access brought about by streaming unlearning requests, we introduce a streaming unlearning paradigm, formalizing the unlearning as a distribution shift problem. We then estimate the altered distribution and propose a novel streaming unlearning algorithm to achieve efficient streaming forgetting without requiring access to the original training data. Theoretical analyses confirm an $O(\sqrt{T} + V_T)$ error bound on the streaming unlearning regret, where $V_T$ represents the cumulative total variation in the optimal solution over $T$ learning rounds. This theoretical guarantee is achieved under mild conditions without the strong restriction of convex loss function. Experiments across various models and datasets validate the performance of our proposed method.
format Preprint
id arxiv_https___arxiv_org_abs_2507_15280
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Machine Unlearning for Streaming Forgetting
Shen, Shaofei
Zhang, Chenhao
Zhao, Yawen
Bialkowski, Alina
Chen, Weitong
Xu, Miao
Machine Learning
Machine unlearning aims to remove knowledge of the specific training data in a well-trained model. Currently, machine unlearning methods typically handle all forgetting data in a single batch, removing the corresponding knowledge all at once upon request. However, in practical scenarios, requests for data removal often arise in a streaming manner rather than in a single batch, leading to reduced efficiency and effectiveness in existing methods. Such challenges of streaming forgetting have not been the focus of much research. In this paper, to address the challenges of performance maintenance, efficiency, and data access brought about by streaming unlearning requests, we introduce a streaming unlearning paradigm, formalizing the unlearning as a distribution shift problem. We then estimate the altered distribution and propose a novel streaming unlearning algorithm to achieve efficient streaming forgetting without requiring access to the original training data. Theoretical analyses confirm an $O(\sqrt{T} + V_T)$ error bound on the streaming unlearning regret, where $V_T$ represents the cumulative total variation in the optimal solution over $T$ learning rounds. This theoretical guarantee is achieved under mild conditions without the strong restriction of convex loss function. Experiments across various models and datasets validate the performance of our proposed method.
title Machine Unlearning for Streaming Forgetting
topic Machine Learning
url https://arxiv.org/abs/2507.15280