3 papers
cs.CV2020
Stochastic Talking Face Generation Using Latent Distribution Matching
Ravindra Yadav, Ashish Sardana, Vinay P Namboodiri +1
The ability to envisage the visual of a talking face based just on hearing a voice is a unique human capability. There have been a number of works that have solved for this ability…
cs.CV2020
Speech Prediction in Silent Videos using Variational Autoencoders
Ravindra Yadav, Ashish Sardana, Vinay P Namboodiri +1
Understanding the relationship between the auditory and visual signals is crucial for many different applications ranging from computer-generated imagery (CGI) and video editing au…
cs.CL2020
Multilogue-Net: A Context Aware RNN for Multi-modal Emotion Detection and Sentiment Analysis in Conversation
Aman Shenoy, Ashish Sardana
Sentiment Analysis and Emotion Detection in conversation is key in several real-world applications, with an increase in modalities available aiding a better understanding of the un…