Large Language Model Federated Learning with Blockchain and Unlearning for Cross-Organizational Collaboration

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: Zuo, Xuhan, Wang, Minghao, Zhu, Tianqing, Yu, Shui, Zhou, Wanlei
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866910750697586688
author Zuo, Xuhan
Wang, Minghao
Zhu, Tianqing
Yu, Shui
Zhou, Wanlei
author_facet Zuo, Xuhan
Wang, Minghao
Zhu, Tianqing
Yu, Shui
Zhou, Wanlei
contents Large language models (LLMs) have transformed the way computers understand and process human language, but using them effectively across different organizations remains still difficult. When organizations work together to improve LLMs, they face several main challenges. First, organizations hesitate to share their valuable data with others. Second, competition between organizations creates trust problems during collaboration. Third, new privacy laws require organizations to be able to delete specific data when requested, which is especially difficult when multiple organizations are learning from shared data. Traditional federated learning approaches do not address these interconnected challenges, particularly in scenarios where participants cannot fully trust each other or the central aggregator. To overcome these limitations, we propose a hybrid blockchain-based federated learning framework that uniquely combines public and private blockchain architectures with multi-agent reinforcement learning. Our framework enables transparent sharing of model update through the public blockchain while protecting sensitive computations in private chains. Each organization operates as an intelligent agent, using Q-learning to optimize its participation strategy and resource allocation, thus aligning individual incentives with collective goals. Notably, we introduce an efficient unlearning mechanism based on Low-Rank Adaptation (LoRA) that enables selective removal of specific data contributions without compromising the model's overall performance. Through extensive experimentation on real-world datasets, we demonstrate that our framework effectively balances privacy protection, trust establishment, and regulatory compliance while maintaining high model performance.
format Preprint
id arxiv_https___arxiv_org_abs_2412_13551
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Large Language Model Federated Learning with Blockchain and Unlearning for Cross-Organizational Collaboration
Zuo, Xuhan
Wang, Minghao
Zhu, Tianqing
Yu, Shui
Zhou, Wanlei
Cryptography and Security
Large language models (LLMs) have transformed the way computers understand and process human language, but using them effectively across different organizations remains still difficult. When organizations work together to improve LLMs, they face several main challenges. First, organizations hesitate to share their valuable data with others. Second, competition between organizations creates trust problems during collaboration. Third, new privacy laws require organizations to be able to delete specific data when requested, which is especially difficult when multiple organizations are learning from shared data. Traditional federated learning approaches do not address these interconnected challenges, particularly in scenarios where participants cannot fully trust each other or the central aggregator. To overcome these limitations, we propose a hybrid blockchain-based federated learning framework that uniquely combines public and private blockchain architectures with multi-agent reinforcement learning. Our framework enables transparent sharing of model update through the public blockchain while protecting sensitive computations in private chains. Each organization operates as an intelligent agent, using Q-learning to optimize its participation strategy and resource allocation, thus aligning individual incentives with collective goals. Notably, we introduce an efficient unlearning mechanism based on Low-Rank Adaptation (LoRA) that enables selective removal of specific data contributions without compromising the model's overall performance. Through extensive experimentation on real-world datasets, we demonstrate that our framework effectively balances privacy protection, trust establishment, and regulatory compliance while maintaining high model performance.
title Large Language Model Federated Learning with Blockchain and Unlearning for Cross-Organizational Collaboration
topic Cryptography and Security
url https://arxiv.org/abs/2412.13551