3 papers
eess.AS2020
Attention Driven Fusion for Multi-Modal Emotion Recognition
Darshana Priyasad, Tharindu Fernando, Simon Denman +2
Deep learning has emerged as a powerful alternative to hand-crafted methods for emotion recognition on combined acoustic and text modalities. Baseline systems model emotion informa…
cs.LG2020
Memory based fusion for multi-modal deep learning
Darshana Priyasad, Tharindu Fernando, Simon Denman +2
The use of multi-modal data for deep machine learning has shown promise when compared to uni-modal approaches with fusion of multi-modal features resulting in improved performance…
eess.AS2020
Temporarily-Aware Context Modelling using Generative Adversarial Networks for Speech Activity Detection
Tharindu Fernando, Sridha Sridharan, Mitchell McLaren +3
This paper presents a novel framework for Speech Activity Detection (SAD). Inspired by the recent success of multi-task learning approaches in the speech processing domain, we prop…