4 papers · 1 filter
ASDnB: Merging Face with Body Cues For Robust Active Speaker Detection
Tiago Roxo, Joana C. Costa, Pedro Inácio +1
State-of-the-art Active Speaker Detection (ASD) approaches mainly use audio and facial features as input. However, the main hypothesis in this paper is that body dynamics is also h…
BIAS: A Body-based Interpretable Active Speaker Approach
Tiago Roxo, Joana C. Costa, Pedro R. M. Inácio +1
State-of-the-art Active Speaker Detection (ASD) approaches heavily rely on audio and facial features to perform, which is not a sustainable approach in wild scenarios. Although the…
How to Squeeze An Explanation Out of Your Model
Tiago Roxo, Joana C. Costa, Pedro R. M. Inácio +1
Deep learning models are widely used nowadays for their reliability in performing various tasks. However, they do not typically provide the reasoning behind their decision, which i…
WASD: A Wilder Active Speaker Detection Dataset
Tiago Roxo, Joana C. Costa, Pedro R. M. Inácio +1
Current Active Speaker Detection (ASD) models achieve great results on AVA-ActiveSpeaker (AVA), using only sound and facial features. Although this approach is applicable in movie…