1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.CV2021★ 1 cited
Localizing Visual Sounds the Hard Way
Honglie Chen, Weidi Xie, Triantafyllos Afouras +3
The objective of this work is to localize sound sources that are visible in a video without using manual annotations. Our key technical contribution is to show that, by training th…
cs.CV2020
VGGSound: A Large-scale Audio-Visual Dataset
Honglie Chen, Weidi Xie, Andrea Vedaldi +1
Our goal is to collect a large-scale audio-visual dataset with low label noise from videos in the wild using computer vision techniques. The resulting dataset can be used for train…
cs.CV2019
AutoCorrect: Deep Inductive Alignment of Noisy Geometric Annotations
Honglie Chen, Weidi Xie, Andrea Vedaldi +1
We propose AutoCorrect, a method to automatically learn object-annotation alignments from a dataset with annotations affected by geometric noise. The method is based on a consisten…