Q-GADMM: Quantized Group ADMM for Communication Efficient Decentralized Machine Learning

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Elgabli, Anis, Park, Jihong, Bedi, Amrit S., Issaid, Chaouki Ben, Bennis, Mehdi, Aggarwal, Vaneet
Format: Preprint
Veröffentlicht: 2019
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866909459154993152
author Elgabli, Anis
Park, Jihong
Bedi, Amrit S.
Issaid, Chaouki Ben
Bennis, Mehdi
Aggarwal, Vaneet
author_facet Elgabli, Anis
Park, Jihong
Bedi, Amrit S.
Issaid, Chaouki Ben
Bennis, Mehdi
Aggarwal, Vaneet
contents In this article, we propose a communication-efficient decentralized machine learning (ML) algorithm, coined quantized group ADMM (Q-GADMM). To reduce the number of communication links, every worker in Q-GADMM communicates only with two neighbors, while updating its model via the group alternating direction method of multipliers (GADMM). Moreover, each worker transmits the quantized difference between its current model and its previously quantized model, thereby decreasing the communication payload size. However, due to the lack of centralized entity in decentralized ML, the spatial sparsity and payload compression may incur error propagation, hindering model training convergence. To overcome this, we develop a novel stochastic quantization method to adaptively adjust model quantization levels and their probabilities, while proving the convergence of Q-GADMM for convex objective functions. Furthermore, to demonstrate the feasibility of Q-GADMM for non-convex and stochastic problems, we propose quantized stochastic GADMM (Q-SGADMM) that incorporates deep neural network architectures and stochastic sampling. Simulation results corroborate that Q-GADMM significantly outperforms GADMM in terms of communication efficiency while achieving the same accuracy and convergence speed for a linear regression task. Similarly, for an image classification task using DNN, Q-SGADMM achieves significantly less total communication cost with identical accuracy and convergence speed compared to its counterpart without quantization, i.e., stochastic GADMM (SGADMM).
format Preprint
id arxiv_https___arxiv_org_abs_1910_10453
institution arXiv
publishDate 2019
record_format arxiv
spellingShingle Q-GADMM: Quantized Group ADMM for Communication Efficient Decentralized Machine Learning
Elgabli, Anis
Park, Jihong
Bedi, Amrit S.
Issaid, Chaouki Ben
Bennis, Mehdi
Aggarwal, Vaneet
Machine Learning
Distributed, Parallel, and Cluster Computing
Information Theory
Networking and Internet Architecture
In this article, we propose a communication-efficient decentralized machine learning (ML) algorithm, coined quantized group ADMM (Q-GADMM). To reduce the number of communication links, every worker in Q-GADMM communicates only with two neighbors, while updating its model via the group alternating direction method of multipliers (GADMM). Moreover, each worker transmits the quantized difference between its current model and its previously quantized model, thereby decreasing the communication payload size. However, due to the lack of centralized entity in decentralized ML, the spatial sparsity and payload compression may incur error propagation, hindering model training convergence. To overcome this, we develop a novel stochastic quantization method to adaptively adjust model quantization levels and their probabilities, while proving the convergence of Q-GADMM for convex objective functions. Furthermore, to demonstrate the feasibility of Q-GADMM for non-convex and stochastic problems, we propose quantized stochastic GADMM (Q-SGADMM) that incorporates deep neural network architectures and stochastic sampling. Simulation results corroborate that Q-GADMM significantly outperforms GADMM in terms of communication efficiency while achieving the same accuracy and convergence speed for a linear regression task. Similarly, for an image classification task using DNN, Q-SGADMM achieves significantly less total communication cost with identical accuracy and convergence speed compared to its counterpart without quantization, i.e., stochastic GADMM (SGADMM).
title Q-GADMM: Quantized Group ADMM for Communication Efficient Decentralized Machine Learning
topic Machine Learning
Distributed, Parallel, and Cluster Computing
Information Theory
Networking and Internet Architecture
url https://arxiv.org/abs/1910.10453