22 citations · 22 across the 1 of their papers we have counts for
2 papers
cs.CV2021
Robust Audio-Visual Instance Discrimination
Pedro Morgado, Ishan Misra, Nuno Vasconcelos
We present a self-supervised learning method to learn audio and video representations. Prior work uses the natural correspondence between audio and video to define a standard cross…
cs.CV2020★ 22 cited
Learning Representations from Audio-Visual Spatial Alignment
Pedro Morgado, Yi Li, Nuno Vasconcelos
We introduce a novel self-supervised pretext task for learning representations from audio-visual content. Prior work on audio-visual representation learning leverages correspondenc…