26 citations · 27 across the 3 of their papers we have counts for
3 papers
cs.SD2022★ 1 cited
Optimizing Bilingual Neural Transducer with Synthetic Code-switching Text Generation
Thien Nguyen, Nathalie Tran, Liuhui Deng +16
Code-switching describes the practice of using more than one language in the same sentence. In this study, we investigate how to optimize a neural transducer based bilingual automa…
eess.AS2020
A Density Ratio Approach to Language Model Fusion in End-To-End Automatic Speech Recognition
Erik McDermott, Hasim Sak, Ehsan Variani
This article describes a density ratio approach to integrating external Language Models (LMs) into end-to-end models for Automatic Speech Recognition (ASR). Applied to a Recurrent…
eess.AS2020★ 26 cited
Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss
Qian Zhang, Han Lu, Hasim Sak +4
In this paper we present an end-to-end speech recognition model with Transformer encoders that can be used in a streaming speech recognition system. Transformer computation blocks…