328 citations · 335 across the 6 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Video Occupancy Models
Manan Tomar, Philippe Hansen-Estruch, Philip Bachman +4
We introduce a new family of video prediction models designed to support downstream control tasks. We call these models Video Occupancy models (VOCs). VOCs operate in a compact lat…
cs.CV2023
Leveraging the Third Dimension in Contrastive Learning
Sumukh Aithal, Anirudh Goyal, Alex Lamb +2
Self-Supervised Learning (SSL) methods operate on unlabeled data to learn robust representations useful for downstream tasks. Most SSL methods rely on augmentations obtained by tra…