Decentralized Stochastic Gradient Langevin Dynamics and Hamiltonian Monte Carlo
arXiv:2007.00590
Abstract
Stochastic gradient Langevin dynamics (SGLD) and stochastic gradient Hamiltonian Monte Carlo (SGHMC) are two popular Markov Chain Monte Carlo (MCMC) algorithms for Bayesian inference that can scale to large datasets, allowing to sample from the posterior distribution of the parameters of a statistical model given the input data and the prior distribution over the model parameters. However, these algorithms do not apply to the decentralized learning setting, when a network of agents are working collaboratively to learn the parameters of a statistical model without sharing their individual data due to privacy reasons or communication constraints. We study two algorithms: Decentralized SGLD (DE-SGLD) and Decentralized SGHMC (DE-SGHMC) which are adaptations of SGLD and SGHMC methods that allow scaleable Bayesian inference in the decentralized setting for large datasets. We show that when the posterior distribution is strongly log-concave and smooth, the iterates of these algorithms converge linearly to a neighborhood of the target distribution in the 2-Wasserstein distance if their parameters are selected appropriately. We illustrate the efficiency of our algorithms on decentralized Bayesian linear regression and Bayesian logistic regression problems.
References in corpus (14)
- Deep Learning: A Bayesian Perspective
- On the Convergence of Stochastic Gradient MCMC Algorithms with High-Order Integrators
- Underdamped Langevin MCMC: A non-asymptotic analysis
- From Averaging to Acceleration, There is Only a Step-size
- Decentralized Bayesian Learning over Graphs
- Robust Distributed Accelerated Stochastic Gradient Methods for Multi-Agent Networks
- An Accelerated Decentralized Stochastic Proximal Algorithm for Finite Sums
- Nonasymptotic estimates for Stochastic Gradient Langevin Dynamics under local conditions in nonconvex optimization
- Variational consensus Monte Carlo
- Differentially Private Accelerated Optimization Algorithms
- Distributed Stochastic Gradient Descent: Nonconvexity, Nonsmoothness, and Convergence to Local Minima
- Nonasymptotic analysis of Stochastic Gradient Hamiltonian Monte Carlo under local conditions for nonconvex optimization
- Decentralized Langevin Dynamics
- Distributed Gradient Methods for Nonconvex Optimization: Local and Global Convergence Guarantees
Cited by in corpus (5)
- Robust Distributed Accelerated Stochastic Gradient Methods for Multi-Agent Networks
- Channel-driven Decentralized Bayesian Federated Learning for Trustworthy Decision Making in D2D Networks
- Decentralized Langevin Dynamics over a Directed Graph
- Heavy-Tail Phenomenon in Decentralized SGD
- A Decentralized Approach to Bayesian Learning