75.5k citations · 83.1k across the 19 of their papers we have counts for
16 papers · 1 filter
Automatic dense annotation of large-vocabulary sign language videos
Liliane Momeni, Hannah Bull, K R Prajwal +3
Recently, sign language researchers have turned to sign language interpreted TV broadcasts, comprising (i) a video of continuous signing and (ii) subtitles corresponding to the aud…
Is an Object-Centric Video Representation Beneficial for Transfer?
Chuhan Zhang, Ankush Gupta, Andrew Zisserman
The objective of this work is to learn an object-centric video representation, with the aim of improving transferability to novel tasks, i.e., tasks different from the pre-training…
Segmenting Moving Objects via an Object-Centric Layered Representation
Junyu Xie, Weidi Xie, Andrew Zisserman
The objective of this paper is a model that is able to discover, track and segment multiple moving objects in a video. We make four contributions: First, we introduce an object-cen…
Audio-Visual Synchronisation in the wild
Honglie Chen, Weidi Xie, Triantafyllos Afouras +3
In this paper, we consider the problem of audio-visual synchronisation applied to videos `in-the-wild' (ie of general classes beyond speech). As a new task, we identify and curate…
Input-level Inductive Biases for 3D Reconstruction
Wang Yifan, Carl Doersch, Relja Arandjelović +2
Much of the recent progress in 3D vision has been driven by the development of specialized architectures that incorporate geometrical inductive biases. In this paper we tackle 3D r…
Interferences in match kernels
Naila Murray, Hervé Jégou, Florent Perronnin +1
We consider the design of an image representation that embeds and aggregates a set of local descriptors into a single vector. Popular representations of this kind include the bag-o…