activity
20182024
most citedLingvo: a Modular and Scalable Framework for Sequence-to-Sequence Modeling

184 citations · 187 across the 10 of their papers we have counts for

collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS20241 cited

ASTRA: Aligning Speech and Text Representations for Asr without Sampling

Neeraj Gaur, Rohan Agrawal, Gary Wang +3

This paper introduces ASTRA, a novel method for improving Automatic Speech Recognition (ASR) through text injection.Unlike prevailing techniques, ASTRA eliminates the need for samp…

eess.AS2022

Accelerating RNN-T Training and Inference Using CTC guidance

Yongqiang Wang, Zhehuai Chen, Chengjian Zheng +3

We propose a novel method to accelerate training and inference process of recurrent neural network transducer (RNN-T) based on the guidance from a co-trained connectionist temporal…

eess.AS2022

Streaming End-to-End Multilingual Speech Recognition with Joint Language Identification

Chao Zhang, Bo Li, Tara Sainath +4

Language identification is critical for many downstream tasks in automatic speech recognition (ASR), and is beneficial to integrate into multilingual end-to-end ASR as an additiona…

eess.AS20221 cited

Unsupervised Data Selection via Discrete Speech Representation for ASR

Zhiyun Lu, Yongqiang Wang, Yu Zhang +3

Self-supervised learning of speech representations has achieved impressive results in improving automatic speech recognition (ASR). In this paper, we show that data selection is im…

eess.AS2018

From Audio to Semantics: Approaches to end-to-end spoken language understanding

Parisa Haghani, Arun Narayanan, Michiel Bacchiani +6

Conventional spoken language understanding systems consist of two main components: an automatic speech recognition module that converts audio to a transcript, and a natural languag…