On parameter estimation with the Wasserstein distance
arXiv:1701.05146
Abstract
Statistical inference can be performed by minimizing, over the parameter space, the Wasserstein distance between model distributions and the empirical distribution of the data. We study asymptotic properties of such minimum Wasserstein distance estimators, complementing results derived by Bassetti, Bodini and Regazzini in 2006. In particular, our results cover the misspecified setting, in which the data-generating process is not assumed to be part of the family of distributions described by the model. Our results are motivated by recent applications of minimum Wasserstein estimators to complex generative models. We discuss some difficulties arising in the approximation of these estimators and illustrate their behavior in several numerical experiments. Two of our examples are taken from the literature on approximate Bayesian computation and have likelihood functions that are not analytically tractable. Two other examples involve misspecified models.
29 pages (+18 pages of appendices), 6 figures. To appear in Information and Inference: A Journal of the IMA. A previous version of this paper contained work on approximate Bayesian computation with the Wasserstein distance, which can now be found at arxiv:1905.03747
References in corpus (7)
- Near-linear time approximation algorithms for optimal transport via Sinkhorn iteration
- Learning in Implicit Generative Models
- The Shortlist Method for Fast Computation of the Earth Mover's Distance and Finding Optimal Solutions to Transportation Problems
- GAN and VAE from an Optimal Transport Point of View
- Massively scalable Sinkhorn distances via the Nyström method
- Optimal transport natural gradient for statistical manifolds with continuous sample space
- Coupling of Particle Filters
Cited by in corpus (10)
- Statistical Aspects of Wasserstein Distances
- Approximate Bayesian computation with the Wasserstein distance
- Learning Generative Models with Sinkhorn Divergences
- GAN and VAE from an Optimal Transport Point of View
- A Tutorial on Deep Latent Variable Models of Natural Language
- Rates of Estimation of Optimal Transport Maps using Plug-in Estimators via Barycentric Projections
- Minimax Distribution Estimation in Wasserstein Distance
- Optimal transport natural gradient for statistical manifolds with continuous sample space
- The energy distance for ensemble and scenario reduction
- Stability of Gibbs Posteriors from the Wasserstein Loss for Bayesian Full Waveform Inversion