102 citations · 111 across the 3 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2020
DiDiSpeech: A Large Scale Mandarin Speech Corpus
Tingwei Guo, Cheng Wen, Dongwei Jiang +8
This paper introduces a new open-sourced Mandarin speech corpus, called DiDiSpeech. It consists of about 800 hours of speech data at 48kHz sampling rate from 6000 speakers and the…
eess.AS2020
Transformer based unsupervised pre-training for acoustic representation learning
Ruixiong Zhang, Haiwei Wu, Wubo Li +3
Recently, a variety of acoustic tasks and related applications arised. For many acoustic tasks, the labeled data size may be limited. To handle this problem, we propose an unsuperv…
eess.AS2020★ 9 cited
A Further Study of Unsupervised Pre-training for Transformer Based Speech Recognition
Dongwei Jiang, Wubo Li, Ruixiong Zhang +5
Building a good speech recognition system usually requires large amounts of transcribed data, which is expensive to collect. To tackle this problem, many unsupervised pre-training…