Multi-consensus Decentralized Accelerated Gradient Descent
arXiv:2005.00797
Abstract
This paper considers the decentralized convex optimization problem, which has a wide range of applications in large-scale machine learning, sensor networks, and control theory. We propose novel algorithms that achieve optimal computation complexity and near optimal communication complexity. Our theoretical results give affirmative answers to the open problem on whether there exists an algorithm that can achieve a communication complexity (nearly) matching the lower bound depending on the global condition number instead of the local one. Furthermore, the linear convergence of our algorithms only depends on the strong convexity of global objective and it does \emph{not} require the local functions to be convex. The design of our methods relies on a novel integration of well-known techniques including Nesterov's acceleration, multi-consensus and gradient-tracking. Empirical studies show the outperformance of our methods for machine learning applications.
References in corpus (1)
Cited by in corpus (17)
- FedML: A Research Library and Benchmark for Federated Machine Learning
- DC-DistADMM: ADMM Algorithm for Constrained Distributed Optimization over Directed Graphs
- Recent theoretical advances in decentralized distributed convex optimization
- Accelerated Gradient Tracking over Time-varying Graphs for Decentralized Optimization
- ADOM: Accelerated Decentralized Optimization Method for Time-Varying Networks
- (Corrected Version) Push-LSVRG-UP: Distributed Stochastic Optimization over Unbalanced Directed Networks with Uncoordinated Triggered Probabilities
- Distributed Saddle-Point Problems Under Similarity
- DeEPCA: Decentralized Exact PCA with Linear Convergence Rate
- Acceleration in Distributed Optimization under Similarity
- PMGT-VR: A decentralized proximal-gradient algorithmic framework with variance reduction
- On the Convergence of Nested Decentralized Gradient Methods with Multiple Consensus and Gradient Steps
- Lower Bounds and Optimal Algorithms for Smooth and Strongly Convex Decentralized Optimization Over Time-Varying Networks
- A Distributed Cubic-Regularized Newton Method for Smooth Convex Optimization over Networks
- An Optimal Algorithm for Strongly Convex Minimization under Affine Constraints
- Near-Optimal Decentralized Algorithms for Saddle Point Problems over Time-Varying Networks
- Parallel and Distributed algorithms for ML problems
- Optimal Gradient Tracking for Decentralized Optimization