paper

Disentangling Video with Independent Prediction

arXiv:1901.05590

Abstract

We propose an unsupervised variational model for disentangling video into independent factors, i.e. each factor's future can be predicted from its past without considering the others. We show that our approach often learns factors which are interpretable as objects in a scene.

Presented at the Learning Disentangled Representations: from Perception to Control workshop at NIPS 2017