On the Convergence of Decentralized Gradient Descent
arXiv:1310.7063
Abstract
Consider the consensus problem of minimizing where each is only known to one individual agent out of a connected network of agents. All the agents shall collaboratively solve this problem and obtain the solution subject to data exchanges restricted to between neighboring agents. Such algorithms avoid the need of a fusion center, offer better network load balance, and improve data privacy. We study the decentralized gradient descent method in which each agent updates its variable , which is a local approximate to the unknown variable , by combining the average of its neighbors' with the negative gradient step . The iteration is where the averaging coefficients form a symmetric doubly stochastic matrix . We analyze the convergence of this iteration and derive its converge rate, assuming that each is proper closed convex and lower bounded, is Lipschitz continuous with constant , and stepsize is fixed. Provided that where , the objective error at the averaged solution, , reduces at a speed of until it reaches . If are further (restricted) strongly convex, then both and each converge to the global minimizer at a linear rate until reaching an -neighborhood of . We also develop an iteration for decentralized basis pursuit and establish its linear convergence to an -neighborhood of the true unknown sparse signal.
Cited by in corpus (15)
- DQM: Decentralized Quadratically Approximated Alternating Direction Method of Multipliers
- Dictionary Learning over Distributed Models
- Collaborative Deep Learning in Fixed Topology Networks
- Proximity Without Consensus in Online Multi-Agent Optimization
- ExtraPush for convex smooth decentralized optimization over directed networks
- A Primal-Dual Quasi-Newton Method for Exact Consensus Optimization
- On the Linear Convergence of Distributed Optimization over Directed Graphs
- Network Newton-Part I: Algorithm and Convergence
- EXTRA: An Exact First-Order Algorithm for Decentralized Consensus Optimization
- Automated Worst-Case Performance Analysis of Decentralized Gradient Descent
- Communication-Efficient Zeroth-Order Distributed Online Optimization: Algorithm, Theory, and Applications
- A Bregman Splitting Algorithm for Distributed Optimization over Networks
- Toward Creating Subsurface Camera
- Graph Balancing for Distributed Subgradient Methods over Directed Graphs
- Differentially Private Distributed Online Learning