97 citations · 296 across the 17 of their papers we have counts for
3 papers · 1 filter
Streaming parallel transducer beam search with fast-slow cascaded encoders
Jay Mahadeokar, Yangyang Shi, Ke Li +5
Streaming ASR with strict latency constraints is required in many speech recognition applications. In order to achieve the required latency, streaming ASR models sacrifice accuracy…
Collaborative Training of Acoustic Encoders for Speech Recognition
Varun Nagaraja, Yangyang Shi, Ganesh Venkatesh +3
On-device speech recognition requires training models of different sizes for deploying on devices with various computational budgets. When building such different models, we can be…
Noisy Training Improves E2E ASR for the Edge
Dilin Wang, Yuan Shangguan, Haichuan Yang +6
Automatic speech recognition (ASR) has become increasingly ubiquitous on modern edge devices. Past work developed streaming End-to-End (E2E) all-neural speech recognizers that can…