55 citations · 256 across the 65 of their papers we have counts for
11 papers · 1 filter
Multimodal Emotion Recognition using Transfer Learning from Speaker Recognition and BERT-based models
Sarala Padi, Seyed Omid Sadjadi, Dinesh Manocha +1
Automatic emotion recognition plays a key role in computer-human interaction as it has the potential to enrich the next-generation artificial intelligence with emotional intelligen…
Binaural Audio Generation via Multi-task Learning
Sijia Li, Shiguang Liu, Dinesh Manocha
We present a learning-based approach for generating binaural audio from mono audio using multi-task learning. Our formulation leverages additional information from two related task…
Improved Speech Emotion Recognition using Transfer Learning and Spectrogram Augmentation
Sarala Padi, Seyed Omid Sadjadi, Dinesh Manocha +1
Automatic speech emotion recognition (SER) is a challenging task that plays a crucial role in natural human-computer interaction. One of the main challenges in SER is data scarcity…
Point-based Acoustic Scattering for Interactive Sound Propagation via Surface Encoding
Hsien-Yu Meng, Zhenyu Tang, Dinesh Manocha
We present a novel geometric deep learning method to compute the acoustic scattering properties of geometric objects. Our learning algorithm uses a point cloud representation of ob…
Sound Synthesis, Propagation, and Rendering: A Survey
Shiguang Liu, Dinesh Manocha
Sound, as a crucial sensory channel, plays a vital role in improving the reality and immersiveness of a virtual environment, following only vision in importance. Sound can provide…
IR-GAN: Room Impulse Response Generator for Far-field Speech Recognition
Anton Ratnarajah, Zhenyu Tang, Dinesh Manocha
We present a Generative Adversarial Network (GAN) based room impulse response generator (IR-GAN) for generating realistic synthetic room impulse responses (RIRs). IR-GAN extracts a…