2 papers
cs.CV2020
ViNet: Pushing the limits of Visual Modality for Audio-Visual Saliency Prediction
Samyak Jain, Pradeep Yarlagadda, Shreyank Jyoti +3
We propose the ViNet architecture for audio-visual saliency prediction. ViNet is a fully convolutional encoder-decoder architecture. The encoder uses visual features from a network…
eess.IV2020
Exploiting Temporal Attention Features for Effective Denoising in Videos
Aryansh Omray, Samyak Jain, Utsav Krishnan +1
Video Denoising is one of the fundamental tasks of any videoprocessing pipeline. It is different from image denoising due to the tem-poral aspects of video frames, and any image de…