most citedA Further Study of Unsupervised Pre-training for Transformer Based Speech Recognition

9 citations · 9 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CL2020

Speech SIMCLR: Combining Contrastive and Reconstruction Objective for Self-supervised Speech Representation Learning

Dongwei Jiang, Wubo Li, Miao Cao +2

Self-supervised visual pretraining has shown significant progress recently. Among those methods, SimCLR greatly advanced the state of the art in self-supervised and semi-supervised…

eess.AS2020

DiDiSpeech: A Large Scale Mandarin Speech Corpus

Tingwei Guo, Cheng Wen, Dongwei Jiang +8

This paper introduces a new open-sourced Mandarin speech corpus, called DiDiSpeech. It consists of about 800 hours of speech data at 48kHz sampling rate from 6000 speakers and the…

eess.AS2020

Transformer based unsupervised pre-training for acoustic representation learning

Ruixiong Zhang, Haiwei Wu, Wubo Li +3

Recently, a variety of acoustic tasks and related applications arised. For many acoustic tasks, the labeled data size may be limited. To handle this problem, we propose an unsuperv…

eess.AS20209 cited

A Further Study of Unsupervised Pre-training for Transformer Based Speech Recognition

Dongwei Jiang, Wubo Li, Ruixiong Zhang +5

Building a good speech recognition system usually requires large amounts of transcribed data, which is expensive to collect. To tackle this problem, many unsupervised pre-training…

cs.SD2019

Cross-task pre-training for on-device acoustic scene classification

Ruixiong Zhang, Wei Zou, Xiangang Li

Acoustic scene classification (ASC) and acoustic event detection (AED) are different but related tasks. Acoustic events can provide useful information for recognizing acoustic scen…