850 citations · 1.5k across the 36 of their papers we have counts for
54 papers
Personalized One-Shot Lipreading for an ALS Patient
Bipasha Sen, Aditya Agarwal, Rudrabha Mukhopadhyay +2
Lipreading or visually recognizing speech from the mouth movements of a speaker is a challenging and mentally taxing task. Unfortunately, multiple medical conditions force people t…
Intelligent Video Editing: Incorporating Modern Talking Face Generation Algorithms in a Video Editor
Anchit Gupta, Faizan Farooq Khan, Rudrabha Mukhopadhyay +2
This paper proposes a video editor based on OpenShot with several state-of-the-art facial video editing algorithms as added functionalities. Our editor provides an easy-to-use inte…
Asking questions on handwritten document collections
Minesh Mathew, Lluis Gomez, Dimosthenis Karatzas +1
This work addresses the problem of Question Answering (QA) on handwritten document collections. Unlike typical QA and Visual Question Answering (VQA) formulations where the answer…
Evaluating Computer Vision Techniques for Urban Mobility on Large-Scale, Unconstrained Roads
Harish Rithish, Raghava Modhugu, Ranjith Reddy +2
Conventional approaches for addressing road safety rely on manual interventions or immobile CCTV infrastructure. Such methods are expensive in enforcing compliance to traffic rules…
More Parameters? No Thanks!
Zeeshan Khan, Kartheek Akella, Vinay P. Namboodiri +1
This work studies the long-standing problems of model capacity and negative interference in multilingual neural machine translation MNMT. We use network pruning techniques and obse…
ICDAR2019 Competition on Scanned Receipt OCR and Information Extraction
Zheng Huang, Kai Chen, Jianhua He +4
Scanned receipts OCR and key information extraction (SROIE) represent the processeses of recognizing text from scanned receipts and extracting key texts from them and save the extr…