activity
20172023
most citedConformer: Convolution-augmented Transformer for Speech Recognition

387 citations · 913 across the 34 of their papers we have counts for

collaborators
Showing eess.ASShow all

23 papers · 1 filter

eess.AS2023★ 4 cited

Efficient Adapters for Giant Speech Models

Nanxin Chen, Izhak Shafran, Yu Zhang +4

Large pre-trained speech models are widely used as the de-facto paradigm, especially in scenarios when there is a limited amount of labeled data available. However, finetuning all…

eess.AS2023

LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus

Yuma Koizumi, Heiga Zen, Shigeki Karita +7

This paper introduces a new speech dataset called ``LibriTTS-R'' designed for text-to-speech (TTS) use. It is derived by applying speech restoration to the LibriTTS corpus, which c…

eess.AS2022

Accelerating RNN-T Training and Inference Using CTC guidance

Yongqiang Wang, Zhehuai Chen, Chengjian Zheng +3

We propose a novel method to accelerate training and inference process of recurrent neural network transducer (RNN-T) based on the guidance from a co-trained connectionist temporal…

eess.AS2022★ 1 cited

Unsupervised Data Selection via Discrete Speech Representation for ASR

Zhiyun Lu, Yongqiang Wang, Yu Zhang +3

Self-supervised learning of speech representations has achieved impressive results in improving automatic speech recognition (ASR). In this paper, we show that data selection is im…

eess.AS2021★ 5 cited

WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis

Nanxin Chen, Yu Zhang, Heiga Zen +4

This paper introduces WaveGrad 2, a non-autoregressive generative model for text-to-speech synthesis. WaveGrad 2 is trained to estimate the gradient of the log conditional density…

eess.AS2021

Multi-Task Learning for End-to-End ASR Word and Utterance Confidence with Deletion Prediction

David Qiu, Yanzhang He, Qiujia Li +3

Confidence scores are very useful for downstream applications of automatic speech recognition (ASR) systems. Recent works have proposed using neural networks to learn word or utter…