8 citations · 13 across the 17 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2020
Self-Supervised MultiModal Versatile Networks
Jean-Baptiste Alayrac, Adrià Recasens, Rosalia Schneider +6
Videos are a rich source of multi-modal supervision. In this work, we learn representations using self-supervision by leveraging three modalities naturally present in videos: visua…
cs.CV2018
Variational Saccading: Efficient Inference for Large Resolution Images
Jason Ramapuram, Maurits Diephuis, Frantzeska Lavda +2
Image classification with deep neural networks is typically restricted to images of small dimensionality such as 224 x 244 in Resnet models [24]. This limitation excludes the 4000…