81 citations · 407 across the 23 of their papers we have counts for
34 papers
arXiVeri: Automatic table verification with GPT
Gyungin Shin, Weidi Xie, Samuel Albanie
Without accurate transcription of numerical data in scientific documents, a scientist cannot draw accurate conclusions. Unfortunately, the process of copying numerical data from on…
Open-vocabulary Semantic Segmentation with Frozen Vision-Language Models
Chaofan Ma, Yuhuan Yang, Yanfeng Wang +2
When trained at a sufficient scale, self-supervised learning has exhibited a notable ability to solve a wide range of visual or language understanding tasks. In this paper, we inve…
A Tri-Layer Plugin to Improve Occluded Detection
Guanqi Zhan, Weidi Xie, Andrew Zisserman
Detecting occluded objects still remains a challenge for state-of-the-art object detectors. The objective of this work is to improve the detection for such objects, and thereby imp…
Sparse in Space and Time: Audio-visual Synchronisation with Trainable Selectors
Vladimir Iashin, Weidi Xie, Esa Rahtu +1
The objective of this paper is audio-visual synchronisation of general videos 'in the wild'. For such videos, the events that may be harnessed for synchronisation cues may be spati…
Turbo Training with Token Dropout
Tengda Han, Weidi Xie, Andrew Zisserman
The objective of this paper is an efficient training method for video tasks. We make three contributions: (1) We propose Turbo training, a simple and versatile training paradigm fo…
A Simple Plugin for Transforming Images to Arbitrary Scales
Qinye Zhou, Ziyi Li, Weidi Xie +3
Existing models on super-resolution often specialized for one scale, fundamentally limiting their use in practical scenarios. In this paper, we aim to develop a general plugin that…