Training Generative Reversible Networks
arXiv:1806.01610
Abstract
Generative models with an encoding component such as autoencoders currently receive great interest. However, training of autoencoders is typically complicated by the need to train a separate encoder and decoder model that have to be enforced to be reciprocal to each other. To overcome this problem, by-design reversible neural networks (RevNets) had been previously used as generative models either directly optimizing the likelihood of the data under the model or using an adversarial approach on the generated data. Here, we instead investigate their performance using an adversary on the latent space in the adversarial autoencoder framework. We investigate the generative performance of RevNets on the CelebA dataset, showing that generative RevNets can generate coherent faces with similar quality as Variational Autoencoders. This first attempt to use RevNets inside the adversarial autoencoder framework slightly underperformed relative to recent advanced generative models using an autoencoder component on CelebA, but this gap may diminish with further optimization of the training setup of generative RevNets. In addition to the experiments on CelebA, we show a proof-of-principle experiment on the MNIST dataset suggesting that adversary-free trained RevNets can discover meaningful latent dimensions without pre-specifying the number of dimensions of the latent sampling distribution. In summary, this study shows that RevNets can be employed in different generative training settings. Source code for this study is at https://github.com/robintibor/generative-reversible
Source code for this study is at https://github.com/robintibor/generative-reversible
References in corpus (14)
- GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
- Self-Attention Generative Adversarial Networks
- NICE: Non-linear Independent Components Estimation
- MMD GAN: Towards Deeper Understanding of Moment Matching Network
- The Cramer Distance as a Solution to Biased Wasserstein Gradients
- The Reversible Residual Network: Backpropagation Without Storing Activations
- Demystifying MMD GANs
- Flow-GAN: Combining Maximum Likelihood and Adversarial Learning in Generative Models
- Geometric GAN
- Sliced-Wasserstein Autoencoder: An Embarrassingly Simple Generative Model
- i-RevNet: Deep Invertible Networks
- Improving GANs Using Optimal Transport
- Comparison of Maximum Likelihood and GAN-based training of Real NVPs
- On the Latent Space of Wasserstein Auto-Encoders