125 citations · 170 across the 11 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2021
HMM-Free Encoder Pre-Training for Streaming RNN Transducer
Lu Huang, Jingyu Sun, Yufeng Tang +4
This work describes an encoder pre-training procedure using frame-wise label to improve the training of streaming recurrent neural network transducer (RNN-T) model. Streaming RNN-T…
eess.AS2020★ 5 cited
Improving RNN transducer with normalized jointer network
Mingkun Huang, Jun Zhang, Meng Cai +5
Recurrent neural transducer (RNN-T) is a promising end-to-end (E2E) model in automatic speech recognition (ASR). It has shown superior performance compared to traditional hybrid AS…
eess.AS2020★ 2 cited
Dynamic latency speech recognition with asynchronous revision
Mingkun Huang, Meng Cai, Jun Zhang +4
In this work we propose an inference technique, asynchronous revision, to unify streaming and non-streaming speech recognition models. Specifically, we achieve dynamic latency with…