2 papers
cs.CV2020
ViNet: Pushing the limits of Visual Modality for Audio-Visual Saliency Prediction
Samyak Jain, Pradeep Yarlagadda, Shreyank Jyoti +3
We propose the ViNet architecture for audio-visual saliency prediction. ViNet is a fully convolutional encoder-decoder architecture. The encoder uses visual features from a network…
cs.CV2018
Expression Empowered ResiDen Network for Facial Action Unit Detection
Shreyank Jyoti, Abhinav Dhall
The paper explores the topic of Facial Action Unit (FAU) detection in the wild. In particular, we are interested in answering the following questions: (1) how useful are residual c…