43 citations · 66 across the 4 of their papers we have counts for
3 papers · 2 filters
Multimodal Fusion with Deep Neural Networks for Audio-Video Emotion Recognition
Juan D. S. Ortega, Mohammed Senoussaoui, Eric Granger +3
This paper presents a novel deep neural network (DNN) for multimodal fusion of audio, video and text modalities for emotion recognition. The proposed DNN architecture has independe…
Min-max Entropy for Weakly Supervised Pointwise Localization
Soufiane Belharbi, Jérôme Rony, Jose Dolz +3
Pointwise localization allows more precise localization and accurate interpretability, compared to bounding box, in applications where objects are highly unstructured such as in me…
Audio-Visual Kinship Verification
Xiaoting Wu, Eric Granger, Xiaoyi Feng
Visual kinship verification entails confirming whether or not two individuals in a given pair of images or videos share a hypothesized kin relation. As a generalized face verificat…