5 citations · 7 across the 2 of their papers we have counts for
2 papers
eess.AS2020★ 5 cited
Improving RNN transducer with normalized jointer network
Mingkun Huang, Jun Zhang, Meng Cai +5
Recurrent neural transducer (RNN-T) is a promising end-to-end (E2E) model in automatic speech recognition (ASR). It has shown superior performance compared to traditional hybrid AS…
eess.AS2020★ 2 cited
Dynamic latency speech recognition with asynchronous revision
Mingkun Huang, Meng Cai, Jun Zhang +4
In this work we propose an inference technique, asynchronous revision, to unify streaming and non-streaming speech recognition models. Specifically, we achieve dynamic latency with…