3 citations · 3 across the 2 of their papers we have counts for
3 papers · 1 filter
Tragic Talkers: A Shakespearean Sound- and Light-Field Dataset for Audio-Visual Machine Learning Research
Davide Berghi, Marco Volino, Philip J. B. Jackson
3D audio-visual production aims to deliver immersive and interactive experiences to the consumer. Yet, faithfully reproducing real-world 3D scenes remains a challenging task. This…
Visually Supervised Speaker Detection and Localization via Microphone Array
Davide Berghi, Adrian Hilton, Philip J. B. Jackson
Active speaker detection (ASD) is a multi-modal task that aims to identify who, if anyone, is speaking from a set of candidates. Current audio-visual approaches for ASD typically r…
Audio-Visual Spatial Aligment Requirements of Central and Peripheral Object Events
Davide Berghi, Hanne Stenzel, Marco Volino +2
Immersive audio-visual perception relies on the spatial integration of both auditory and visual information which are heterogeneous sensing modalities with different fields of rece…