Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Communication-Efficient Decentralized Learning with Sparsification and Adaptive Peer Selection

Feb 22, 2020

Zhenheng Tang, Shaohuai Shi, Xiaowen Chu

Figure 1 for Communication-Efficient Decentralized Learning with Sparsification and Adaptive Peer Selection

Figure 2 for Communication-Efficient Decentralized Learning with Sparsification and Adaptive Peer Selection

Figure 3 for Communication-Efficient Decentralized Learning with Sparsification and Adaptive Peer Selection

Figure 4 for Communication-Efficient Decentralized Learning with Sparsification and Adaptive Peer Selection

Share this with someone who'll enjoy it:

Abstract:Distributed learning techniques such as federated learning have enabled multiple workers to train machine learning models together to reduce the overall training time. However, current distributed training algorithms (centralized or decentralized) suffer from the communication bottleneck on multiple low-bandwidth workers (also on the server under the centralized architecture). Although decentralized algorithms generally have lower communication complexity than the centralized counterpart, they still suffer from the communication bottleneck for workers with low network bandwidth. To deal with the communication problem while being able to preserve the convergence performance, we introduce a novel decentralized training algorithm with the following key features: 1) It does not require a parameter server to maintain the model during training, which avoids the communication pressure on any single peer. 2) Each worker only needs to communicate with a single peer at each communication round with a highly compressed model, which can significantly reduce the communication traffic on the worker. We theoretically prove that our sparsification algorithm still preserves convergence properties. 3) Each worker dynamically selects its peer at different communication rounds to better utilize the bandwidth resources. We conduct experiments with convolutional neural networks on 32 workers to verify the effectiveness of our proposed algorithm compared to seven existing methods. Experimental results show that our algorithm significantly reduces the communication traffic and generally select relatively high bandwidth peers.

View paper on

Share this with someone who'll enjoy it:

Title:Communication-Efficient Decentralized Learning with Sparsification and Adaptive Peer Selection

Paper and Code