26 citations · 26 across the 1 of their papers we have counts for
3 papers
cs.SD2020
Transformer Transducer: One Model Unifying Streaming and Non-streaming Speech Recognition
Anshuman Tripathi, Jaeyoung Kim, Qian Zhang +2
In this paper we present a Transformer-Transducer model architecture and a training technique to unify streaming and non-streaming speech recognition models into one model. The mod…
eess.AS2020★ 26 cited
Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss
Qian Zhang, Han Lu, Hasim Sak +4
In this paper we present an end-to-end speech recognition model with Transformer encoders that can be used in a streaming speech recognition system. Transformer computation blocks…
cs.CL2018
Toward domain-invariant speech recognition via large scale training
Arun Narayanan, Ananya Misra, Khe Chai Sim +6
Current state-of-the-art automatic speech recognition systems are trained to work in specific `domains', defined based on factors like application, sampling rate and codec. When su…