3 papers
cs.CV2024
ASDnB: Merging Face with Body Cues For Robust Active Speaker Detection
Tiago Roxo, Joana C. Costa, Pedro Inácio +1
State-of-the-art Active Speaker Detection (ASD) approaches mainly use audio and facial features as input. However, the main hypothesis in this paper is that body dynamics is also h…
cs.CV2024
BIAS: A Body-based Interpretable Active Speaker Approach
Tiago Roxo, Joana C. Costa, Pedro R. M. Inácio +1
State-of-the-art Active Speaker Detection (ASD) approaches heavily rely on audio and facial features to perform, which is not a sustainable approach in wild scenarios. Although the…
cs.CV2024
How to Squeeze An Explanation Out of Your Model
Tiago Roxo, Joana C. Costa, Pedro R. M. Inácio +1
Deep learning models are widely used nowadays for their reliability in performing various tasks. However, they do not typically provide the reasoning behind their decision, which i…