activity
20162024
most citedA Lip Sync Expert Is All You Need for Speech to Lip Generation In The Wild

850 citations · 1.5k across the 40 of their papers we have counts for

collaborators
Showing 2021Show all

7 papers · 1 filter

cs.CV2021★ 1 cited

Personalized One-Shot Lipreading for an ALS Patient

Bipasha Sen, Aditya Agarwal, Rudrabha Mukhopadhyay +2

Lipreading or visually recognizing speech from the mouth movements of a speaker is a challenging and mentally taxing task. Unfortunately, multiple medical conditions force people t…

cs.CV2021★ 7 cited

Intelligent Video Editing: Incorporating Modern Talking Face Generation Algorithms in a Video Editor

Anchit Gupta, Faizan Farooq Khan, Rudrabha Mukhopadhyay +2

This paper proposes a video editor based on OpenShot with several state-of-the-art facial video editing algorithms as added functionalities. Our editor provides an easy-to-use inte…

cs.CV2021★ 12 cited

Asking questions on handwritten document collections

Minesh Mathew, Lluis Gomez, Dimosthenis Karatzas +1

This work addresses the problem of Question Answering (QA) on handwritten document collections. Unlike typical QA and Visual Question Answering (VQA) formulations where the answer…

cs.CV2021

Ego4D: Around the World in 3,000 Hours of Egocentric Video

Kristen Grauman, Andrew Westbury, Eugene Byrne +82

We introduce Ego4D, a massive-scale egocentric video dataset and benchmark suite. It offers 3,670 hours of daily-life activity video spanning hundreds of scenarios (household, outd…

cs.CV2021★ 2 cited

Evaluating Computer Vision Techniques for Urban Mobility on Large-Scale, Unconstrained Roads

Harish Rithish, Raghava Modhugu, Ranjith Reddy +2

Conventional approaches for addressing road safety rely on manual interventions or immobile CCTV infrastructure. Such methods are expensive in enforcing compliance to traffic rules…

cs.CL2021

More Parameters? No Thanks!

Zeeshan Khan, Kartheek Akella, Vinay P. Namboodiri +1

This work studies the long-standing problems of model capacity and negative interference in multilingual neural machine translation MNMT. We use network pruning techniques and obse…