FedGA: Federated Learning with Gradient Alignment for Error Asymmetry Mitigation

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Xiao, Chenguang, Zuo, Zheming, Wang, Shuo
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909437897211904
author Xiao, Chenguang
Zuo, Zheming
Wang, Shuo
author_facet Xiao, Chenguang
Zuo, Zheming
Wang, Shuo
contents Federated learning (FL) triggers intra-client and inter-client class imbalance, with the latter compared to the former leading to biased client updates and thus deteriorating the distributed models. Such a bias is exacerbated during the server aggregation phase and has yet to be effectively addressed by conventional re-balancing methods. To this end, different from the off-the-shelf label or loss-based approaches, we propose a gradient alignment (GA)-informed FL method, dubbed as FedGA, where the importance of error asymmetry (EA) in bias is observed and its linkage to the gradient of the loss to raw logits is explored. Concretely, GA, implemented by label calibration during the model backpropagation process, prevents catastrophic forgetting of rate and missing classes, hence boosting model convergence and accuracy. Experimental results on five benchmark datasets demonstrate that GA outperforms the pioneering counterpart FedAvg and its four variants in minimizing EA and updating bias, and accordingly yielding higher F1 score and accuracy margins when the Dirichlet distribution sampling factor $α$ increases. The code and more details are available at \url{https://anonymous.4open.science/r/FedGA-B052/README.md}.
format Preprint
id arxiv_https___arxiv_org_abs_2412_16582
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle FedGA: Federated Learning with Gradient Alignment for Error Asymmetry Mitigation
Xiao, Chenguang
Zuo, Zheming
Wang, Shuo
Machine Learning
Cryptography and Security
Federated learning (FL) triggers intra-client and inter-client class imbalance, with the latter compared to the former leading to biased client updates and thus deteriorating the distributed models. Such a bias is exacerbated during the server aggregation phase and has yet to be effectively addressed by conventional re-balancing methods. To this end, different from the off-the-shelf label or loss-based approaches, we propose a gradient alignment (GA)-informed FL method, dubbed as FedGA, where the importance of error asymmetry (EA) in bias is observed and its linkage to the gradient of the loss to raw logits is explored. Concretely, GA, implemented by label calibration during the model backpropagation process, prevents catastrophic forgetting of rate and missing classes, hence boosting model convergence and accuracy. Experimental results on five benchmark datasets demonstrate that GA outperforms the pioneering counterpart FedAvg and its four variants in minimizing EA and updating bias, and accordingly yielding higher F1 score and accuracy margins when the Dirichlet distribution sampling factor $α$ increases. The code and more details are available at \url{https://anonymous.4open.science/r/FedGA-B052/README.md}.
title FedGA: Federated Learning with Gradient Alignment for Error Asymmetry Mitigation
topic Machine Learning
Cryptography and Security
url https://arxiv.org/abs/2412.16582