5 citations · 7 across the 3 of their papers we have counts for
5 papers
Improved Language Identification Through Cross-Lingual Self-Supervised Learning
Andros Tjandra, Diptanu Gon Choudhury, Frank Zhang +6
Language identification greatly impacts the success of downstream tasks such as automatic speech recognition. Recently, self-supervised speech representations learned by wav2vec 2.…
A Multi-View Approach To Audio-Visual Speaker Verification
Leda Sarı, Kritika Singh, Jiatong Zhou +3
Although speaker verification has conventionally been an audio-only task, some practical applications provide both audio and visual streams of input. In these cases, the visual str…
Large scale weakly and semi-supervised learning for low-resource video ASR
Kritika Singh, Vimal Manohar, Alex Xiao +7
Many semi- and weakly-supervised approaches have been investigated for overcoming the labeling cost of building high quality speech recognition systems. On the challenging task of…
Training ASR models by Generation of Contextual Information
Kritika Singh, Dmytro Okhonko, Jun Liu +8
Supervised ASR models have reached unprecedented levels of accuracy, thanks in part to ever-increasing amounts of labelled training data. However, in many applications and locales,…
Multilingual Graphemic Hybrid ASR with Massive Data Augmentation
Chunxi Liu, Qiaochu Zhang, Xiaohui Zhang +3
Towards developing high-performing ASR for low-resource languages, approaches to address the lack of resources are to make use of data from multiple languages, and to augment the t…