1 paper · 1 filter
T. Barak, Y. Loewenstein
Video prediction models often combine three components: an encoder from pixel space to a small latent space, a latent space prediction model, and a generative model back to pixel s…