256 citations · 404 across the 19 of their papers we have counts for
4 papers · 1 filter
USM-Lite: Quantization and Sparsity Aware Fine-tuning for Speech Recognition with Universal Speech Models
Shaojin Ding, David Qiu, David Rim +10
End-to-end automatic speech recognition (ASR) models have seen revolutionary quality gains with the recent development of large-scale universal speech models (USM). However, deploy…
Massive End-to-end Models for Short Search Queries
Weiran Wang, Rohit Prabhavalkar, Dongseong Hwang +11
In this work, we investigate two popular end-to-end automatic speech recognition (ASR) models, namely Connectionist Temporal Classification (CTC) and RNN-Transducer (RNN-T), for of…
Towards Word-Level End-to-End Neural Speaker Diarization with Auxiliary Network
Yiling Huang, Weiran Wang, Guanlong Zhao +3
While standard speaker diarization attempts to answer the question "who spoken when", most of relevant applications in reality are more interested in determining "who spoken what".…
JEIT: Joint End-to-End Model and Internal Language Model Training for Speech Recognition
Zhong Meng, Weiran Wang, Rohit Prabhavalkar +7
We propose JEIT, a joint end-to-end (E2E) model and internal language model (ILM) training method to inject large-scale unpaired text into ILM during E2E training which improves ra…