65 citations · 65 across the 2 of their papers we have counts for
4 papers
Face, Body, Voice: Video Person-Clustering with Multiple Modalities
Andrew Brown, Vicky Kalogeiton, Andrew Zisserman
The objective of this work is person-clustering in videos -- grouping characters according to their identity. Previous methods focus on the narrower task of face-clustering, and fo…
Automated Video Labelling: Identifying Faces by Corroborative Evidence
Andrew Brown, Ernesto Coto, Andrew Zisserman
We present a method for automatically labelling all faces in video archives, such as TV broadcasts, by combining multiple evidence sources and multiple modalities (visual and audio…
VoxSRC 2020: The Second VoxCeleb Speaker Recognition Challenge
Arsha Nagrani, Joon Son Chung, Jaesung Huh +6
We held the second installment of the VoxCeleb Speaker Recognition Challenge in conjunction with Interspeech 2020. The goal of this challenge was to assess how well current speaker…
4-Connected Shift Residual Networks
Andrew Brown, Pascal Mettes, Marcel Worring
The shift operation was recently introduced as an alternative to spatial convolutions. The operation moves subsets of activations horizontally and/or vertically. Spatial convolutio…