416 citations · 626 across the 7 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2022★ 3 cited
TESSP: Text-Enhanced Self-Supervised Speech Pre-training
Zhuoyuan Yao, Shuo Ren, Sanyuan Chen +3
Self-supervised speech pre-training empowers the model with the contextual structure inherent in the speech signal while self-supervised text pre-training empowers the model with l…
cs.SD2022
Speech Pre-training with Acoustic Piece
Shuo Ren, Shujie Liu, Yu Wu +2
Previous speech pre-training methods, such as wav2vec2.0 and HuBERT, pre-train a Transformer encoder to learn deep representations from audio data, with objectives predicting eithe…
cs.SD2021
Optimizing Alignment of Speech and Language Latent Spaces for End-to-End Speech Recognition and Understanding
Wei Wang, Shuo Ren, Yao Qian +4
The advances in attention-based encoder-decoder (AED) networks have brought great progress to end-to-end (E2E) automatic speech recognition (ASR). One way to further improve the pe…