Benchmarking Federated Machine Unlearning methods for Tabular Data

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Xiao, Chenguang, Ghosh, Abhirup, Wu, Han, Wang, Shuo, van Thiel, Diederick
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866910900972158976
author Xiao, Chenguang
Ghosh, Abhirup
Wu, Han
Wang, Shuo
van Thiel, Diederick
author_facet Xiao, Chenguang
Ghosh, Abhirup
Wu, Han
Wang, Shuo
van Thiel, Diederick
contents Machine unlearning, which enables a model to forget specific data upon request, is increasingly relevant in the era of privacy-centric machine learning, particularly within federated learning (FL) environments. This paper presents a pioneering study on benchmarking machine unlearning methods within a federated setting for tabular data, addressing the unique challenges posed by cross-silo FL where data privacy and communication efficiency are paramount. We explore unlearning at the feature and instance levels, employing both machine learning, random forest and logistic regression models. Our methodology benchmarks various unlearning algorithms, including fine-tuning and gradient-based approaches, across multiple datasets, with metrics focused on fidelity, certifiability, and computational efficiency. Experiments demonstrate that while fidelity remains high across methods, tree-based models excel in certifiability, ensuring exact unlearning, whereas gradient-based methods show improved computational efficiency. This study provides critical insights into the design and selection of unlearning algorithms tailored to the FL environment, offering a foundation for further research in privacy-preserving machine learning.
format Preprint
id arxiv_https___arxiv_org_abs_2504_00921
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Benchmarking Federated Machine Unlearning methods for Tabular Data
Xiao, Chenguang
Ghosh, Abhirup
Wu, Han
Wang, Shuo
van Thiel, Diederick
Machine Learning
Machine unlearning, which enables a model to forget specific data upon request, is increasingly relevant in the era of privacy-centric machine learning, particularly within federated learning (FL) environments. This paper presents a pioneering study on benchmarking machine unlearning methods within a federated setting for tabular data, addressing the unique challenges posed by cross-silo FL where data privacy and communication efficiency are paramount. We explore unlearning at the feature and instance levels, employing both machine learning, random forest and logistic regression models. Our methodology benchmarks various unlearning algorithms, including fine-tuning and gradient-based approaches, across multiple datasets, with metrics focused on fidelity, certifiability, and computational efficiency. Experiments demonstrate that while fidelity remains high across methods, tree-based models excel in certifiability, ensuring exact unlearning, whereas gradient-based methods show improved computational efficiency. This study provides critical insights into the design and selection of unlearning algorithms tailored to the FL environment, offering a foundation for further research in privacy-preserving machine learning.
title Benchmarking Federated Machine Unlearning methods for Tabular Data
topic Machine Learning
url https://arxiv.org/abs/2504.00921