Federated Deconfounding and Debiasing Learning for Out-of-Distribution Generalization

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Qi, Zhuang, Zhou, Sijin, Meng, Lei, Hu, Han, Yu, Han, Meng, Xiangxu
Format: Preprint
Published: 2025
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913831226179584
author Qi, Zhuang
Zhou, Sijin
Meng, Lei
Hu, Han
Yu, Han
Meng, Xiangxu
author_facet Qi, Zhuang
Zhou, Sijin
Meng, Lei
Hu, Han
Yu, Han
Meng, Xiangxu
contents Attribute bias in federated learning (FL) typically leads local models to optimize inconsistently due to the learning of non-causal associations, resulting degraded performance. Existing methods either use data augmentation for increasing sample diversity or knowledge distillation for learning invariant representations to address this problem. However, they lack a comprehensive analysis of the inference paths, and the interference from confounding factors limits their performance. To address these limitations, we propose the \underline{Fed}erated \underline{D}econfounding and \underline{D}ebiasing \underline{L}earning (FedDDL) method. It constructs a structured causal graph to analyze the model inference process, and performs backdoor adjustment to eliminate confounding paths. Specifically, we design an intra-client deconfounding learning module for computer vision tasks to decouple background and objects, generating counterfactual samples that establish a connection between the background and any label, which stops the model from using the background to infer the label. Moreover, we design an inter-client debiasing learning module to construct causal prototypes to reduce the proportion of the background in prototype components. Notably, it bridges the gap between heterogeneous representations via causal prototypical regularization. Extensive experiments on 2 benchmarking datasets demonstrate that \methodname{} significantly enhances the model capability to focus on main objects in unseen data, leading to 4.5\% higher Top-1 Accuracy on average over 9 state-of-the-art existing methods.
format Preprint
id arxiv_https___arxiv_org_abs_2505_04979
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Federated Deconfounding and Debiasing Learning for Out-of-Distribution Generalization
Qi, Zhuang
Zhou, Sijin
Meng, Lei
Hu, Han
Yu, Han
Meng, Xiangxu
Computer Vision and Pattern Recognition
Attribute bias in federated learning (FL) typically leads local models to optimize inconsistently due to the learning of non-causal associations, resulting degraded performance. Existing methods either use data augmentation for increasing sample diversity or knowledge distillation for learning invariant representations to address this problem. However, they lack a comprehensive analysis of the inference paths, and the interference from confounding factors limits their performance. To address these limitations, we propose the \underline{Fed}erated \underline{D}econfounding and \underline{D}ebiasing \underline{L}earning (FedDDL) method. It constructs a structured causal graph to analyze the model inference process, and performs backdoor adjustment to eliminate confounding paths. Specifically, we design an intra-client deconfounding learning module for computer vision tasks to decouple background and objects, generating counterfactual samples that establish a connection between the background and any label, which stops the model from using the background to infer the label. Moreover, we design an inter-client debiasing learning module to construct causal prototypes to reduce the proportion of the background in prototype components. Notably, it bridges the gap between heterogeneous representations via causal prototypical regularization. Extensive experiments on 2 benchmarking datasets demonstrate that \methodname{} significantly enhances the model capability to focus on main objects in unseen data, leading to 4.5\% higher Top-1 Accuracy on average over 9 state-of-the-art existing methods.
title Federated Deconfounding and Debiasing Learning for Out-of-Distribution Generalization
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2505.04979