850 citations · 997 across the 5 of their papers we have counts for
6 papers
Visual Keyword Spotting with Attention
K R Prajwal, Liliane Momeni, Triantafyllos Afouras +1
In this paper, we consider the task of spotting spoken keywords in silent video sequences -- also known as visual keyword spotting. To this end, we investigate Transformer-based mo…
Visual Speech Enhancement Without A Real Visual Stream
Sindhu B Hegde, K R Prajwal, Rudrabha Mukhopadhyay +2
In this work, we re-think the task of speech enhancement in unconstrained real-world environments. Current state-of-the-art methods use only the audio stream and are limited in the…
A Lip Sync Expert Is All You Need for Speech to Lip Generation In The Wild
K R Prajwal, Rudrabha Mukhopadhyay, Vinay Namboodiri +1
In this work, we investigate the problem of lip-syncing a talking face video of an arbitrary identity to match a target speech segment. Current works excel at producing accurate li…
Learning Individual Speaking Styles for Accurate Lip to Speech Synthesis
K R Prajwal, Rudrabha Mukhopadhyay, Vinay Namboodiri +1
Humans involuntarily tend to infer parts of the conversation from lip movements when the speech is absent or corrupted by external noise. In this work, we explore the task of lip t…
Towards Automatic Face-to-Face Translation
Prajwal K R, Rudrabha Mukhopadhyay, Jerin Philip +3
In light of the recent breakthroughs in automatic machine translation systems, we propose a novel approach that we term as "Face-to-Face Translation". As today's digital communicat…
DRUNET: A Dilated-Residual U-Net Deep Learning Network to Digitally Stain Optic Nerve Head Tissues in Optical Coherence Tomography Images
Sripad Krishna Devalla, Prajwal K. Renukanand, Bharathwaj K. Sreedhar +8
Given that the neural and connective tissues of the optic nerve head (ONH) exhibit complex morphological changes with the development and progression of glaucoma, their simultaneou…