activity
20192021
most citedA Multi-View Approach To Audio-Visual Speaker Verification

5 citations · 7 across the 3 of their papers we have counts for

collaborators

5 papers

cs.CL20212 cited

Improved Language Identification Through Cross-Lingual Self-Supervised Learning

Andros Tjandra, Diptanu Gon Choudhury, Frank Zhang +6

Language identification greatly impacts the success of downstream tasks such as automatic speech recognition. Recently, self-supervised speech representations learned by wav2vec 2.…

cs.SD20215 cited

A Multi-View Approach To Audio-Visual Speaker Verification

Leda Sarı, Kritika Singh, Jiatong Zhou +3

Although speaker verification has conventionally been an audio-only task, some practical applications provide both audio and visual streams of input. In these cases, the visual str…

eess.AS2020

Large scale weakly and semi-supervised learning for low-resource video ASR

Kritika Singh, Vimal Manohar, Alex Xiao +7

Many semi- and weakly-supervised approaches have been investigated for overcoming the labeling cost of building high quality speech recognition systems. On the challenging task of…

cs.CL2019

Training ASR models by Generation of Contextual Information

Kritika Singh, Dmytro Okhonko, Jun Liu +8

Supervised ASR models have reached unprecedented levels of accuracy, thanks in part to ever-increasing amounts of labelled training data. However, in many applications and locales,…

eess.AS2019

Multilingual Graphemic Hybrid ASR with Massive Data Augmentation

Chunxi Liu, Qiaochu Zhang, Xiaohui Zhang +3

Towards developing high-performing ASR for low-resource languages, approaches to address the lack of resources are to make use of data from multiple languages, and to augment the t…