activity
20172023
most citedMemory-augmented Dense Predictive Coding for Video Representation Learning

81 citations · 407 across the 23 of their papers we have counts for

collaborators

34 papers

cs.CL2023

arXiVeri: Automatic table verification with GPT

Gyungin Shin, Weidi Xie, Samuel Albanie

Without accurate transcription of numerical data in scientific documents, a scientist cannot draw accurate conclusions. Unfortunately, the process of copying numerical data from on…

cs.CV202210 cited

Open-vocabulary Semantic Segmentation with Frozen Vision-Language Models

Chaofan Ma, Yuhuan Yang, Yanfeng Wang +2

When trained at a sufficient scale, self-supervised learning has exhibited a notable ability to solve a wide range of visual or language understanding tasks. In this paper, we inve…

cs.CV20223 cited

A Tri-Layer Plugin to Improve Occluded Detection

Guanqi Zhan, Weidi Xie, Andrew Zisserman

Detecting occluded objects still remains a challenge for state-of-the-art object detectors. The objective of this work is to improve the detection for such objects, and thereby imp…

cs.CV20222 cited

Sparse in Space and Time: Audio-visual Synchronisation with Trainable Selectors

Vladimir Iashin, Weidi Xie, Esa Rahtu +1

The objective of this paper is audio-visual synchronisation of general videos 'in the wild'. For such videos, the events that may be harnessed for synchronisation cues may be spati…

cs.CV20224 cited

Turbo Training with Token Dropout

Tengda Han, Weidi Xie, Andrew Zisserman

The objective of this paper is an efficient training method for video tasks. We make three contributions: (1) We propose Turbo training, a simple and versatile training paradigm fo…

cs.CV2022

A Simple Plugin for Transforming Images to Arbitrary Scales

Qinye Zhou, Ziyi Li, Weidi Xie +3

Existing models on super-resolution often specialized for one scale, fundamentally limiting their use in practical scenarios. In this paper, we aim to develop a general plugin that…