Learning Arbitrary Statistical Mixtures of Discrete Distributions
arXiv:1504.02526
Abstract
We study the problem of learning from unlabeled samples very general statistical mixture models on large finite sets. Specifically, the model to be learned, , is a probability distribution over probability distributions , where each such is a probability distribution over . When we sample from , we do not observe directly, but only indirectly and in very noisy fashion, by sampling from repeatedly, independently times from the distribution . The problem is to infer to high accuracy in transportation (earthmover) distance. We give the first efficient algorithms for learning this mixture model without making any restricting assumptions on the structure of the distribution . We bound the quality of the solution as a function of the size of the samples and the number of samples used. Our model and results have applications to a variety of unsupervised learning scenarios, including learning topic models and collaborative filtering.
23 pages. Preliminary version in the Proceeding of the 47th ACM Symposium on the Theory of Computing (STOC15)
References in corpus (6)
- Probabilistic Latent Semantic Analysis
- A Spectral Algorithm for Latent Dirichlet Allocation
- Polynomial Learning of Distribution Families
- Learning mixtures of separated nonspherical Gaussians
- Settling the Polynomial Learnability of Mixtures of Gaussians
- PAC Learning Mixtures of Axis-Aligned Gaussians with No Separation Assumption