Direct Optimization through for Discrete Variational Auto-Encoder
arXiv:1806.02867
Abstract
Reparameterization of variational auto-encoders with continuous random variables is an effective method for reducing the variance of their gradient estimates. In the discrete case, one can perform reparametrization using the Gumbel-Max trick, but the resulting objective relies on an operation and is non-differentiable. In contrast to previous works which resort to softmax-based relaxations, we propose to optimize it directly by applying the direct loss minimization approach. Our proposal extends naturally to structured discrete latent variable models when evaluating the operation is tractable. We demonstrate empirically the effectiveness of the direct loss minimization technique in variational autoencoders with both unstructured and structured discrete latent variables.
Accepted by Neural Information Processing Systems (NeurIPS 2019)
References in corpus (11)
- Fashion-MNIST: a Novel Image Dataset for Benchmarking Machine Learning Algorithms
- The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables
- Junction Tree Variational Autoencoder for Molecular Graph Generation
- Learning Latent Permutations with Gumbel-Sinkhorn Networks
- Grammar Variational Autoencoder
- Learning to Compose Words into Sentences with Reinforcement Learning
- Differentiable Perturb-and-Parse: Semi-Supervised Parsing with a Structured Variational Autoencoder
- DVAE++: Discrete Variational Autoencoders with Overlapping Transformations
- ARM: Augment-REINFORCE-Merge Gradient for Stochastic Binary Networks
- DVAE#: Discrete Variational Autoencoders with Relaxed Boltzmann Priors
- Local Expectation Gradients for Doubly Stochastic Variational Inference