30 citations · 48 across the 3 of their papers we have counts for
3 papers
cs.CL2021★ 6 cited
MixSpeech: Data Augmentation for Low-resource Automatic Speech Recognition
Linghui Meng, Jin Xu, Xu Tan +3
In this paper, we propose MixSpeech, a simple yet effective data augmentation method based on mixup for automatic speech recognition (ASR). MixSpeech trains an ASR model by taking…
eess.AS2020★ 12 cited
LRSpeech: Extremely Low-Resource Speech Synthesis and Recognition
Jin Xu, Xu Tan, Yi Ren +4
Speech synthesis (text to speech, TTS) and recognition (automatic speech recognition, ASR) are important speech tasks, and require a large amount of text and speech pairs for model…
eess.AS2020★ 30 cited
MultiSpeech: Multi-Speaker Text to Speech with Transformer
Mingjian Chen, Xu Tan, Yi Ren +5
Transformer-based text to speech (TTS) model (e.g., Transformer TTS~\cite{li2019neural}, FastSpeech~\cite{ren2019fastspeech}) has shown the advantages of training and inference eff…