66 citations · 100 across the 3 of their papers we have counts for
3 papers
Benchmarking LF-MMI, CTC and RNN-T Criteria for Streaming ASR
Xiaohui Zhang, Frank Zhang, Chunxi Liu +8
In this work, to measure the accuracy and efficiency for a latency-controlled streaming automatic speech recognition (ASR) application, we perform comprehensive evaluations on thre…
RNN-T For Latency Controlled ASR With Improved Beam Search
Mahaveer Jain, Kjell Schubert, Jay Mahadeokar +5
Neural transducer-based systems such as RNN Transducers (RNN-T) for automatic speech recognition (ASR) blend the individual components of a traditional hybrid ASR systems (acoustic…
Transformer-Transducer: End-to-End Speech Recognition with Self-Attention
Ching-Feng Yeh, Jay Mahadeokar, Kaustubh Kalgaonkar +6
We explore options to use Transformer networks in neural transducer for end-to-end speech recognition. Transformer networks use self-attention for sequence modeling and comes with…