activity
20142023
most citedVery Deep Convolutional Networks for Large-Scale Image Recognition

75.5k citations · 83.1k across the 19 of their papers we have counts for

collaborators
Showing cs.CVShow all

16 papers · 1 filter

cs.CV20221 cited

Automatic dense annotation of large-vocabulary sign language videos

Liliane Momeni, Hannah Bull, K R Prajwal +3

Recently, sign language researchers have turned to sign language interpreted TV broadcasts, comprising (i) a video of continuous signing and (ii) subtitles corresponding to the aud…

cs.CV2022

Is an Object-Centric Video Representation Beneficial for Transfer?

Chuhan Zhang, Ankush Gupta, Andrew Zisserman

The objective of this work is to learn an object-centric video representation, with the aim of improving transferability to novel tasks, i.e., tasks different from the pre-training…

cs.CV202213 cited

Segmenting Moving Objects via an Object-Centric Layered Representation

Junyu Xie, Weidi Xie, Andrew Zisserman

The objective of this paper is a model that is able to discover, track and segment multiple moving objects in a video. We make four contributions: First, we introduce an object-cen…

cs.CV2021

Audio-Visual Synchronisation in the wild

Honglie Chen, Weidi Xie, Triantafyllos Afouras +3

In this paper, we consider the problem of audio-visual synchronisation applied to videos `in-the-wild' (ie of general classes beyond speech). As a new task, we identify and curate…

cs.CV2021

Input-level Inductive Biases for 3D Reconstruction

Wang Yifan, Carl Doersch, Relja Arandjelović +2

Much of the recent progress in 3D vision has been driven by the development of specialized architectures that incorporate geometrical inductive biases. In this paper we tackle 3D r…

cs.CV2016

Interferences in match kernels

Naila Murray, Hervé Jégou, Florent Perronnin +1

We consider the design of an image representation that embeds and aggregates a set of local descriptors into a single vector. Popular representations of this kind include the bag-o…