30 citations · 55 across the 4 of their papers we have counts for
5 papers
LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT
Rui Wang, Qibing Bai, Junyi Ao +6
Self-supervised speech representation learning has shown promising results in various speech processing tasks. However, the pre-trained models, e.g., HuBERT, are storage-intensive…
SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
Junyi Ao, Rui Wang, Long Zhou +11
Motivated by the success of T5 (Text-To-Text Transfer Transformer) in pre-trained natural language processing models, we propose a unified-modal SpeechT5 framework that explores th…
Multi-View Self-Attention Based Transformer for Speaker Recognition
Rui Wang, Junyi Ao, Long Zhou +5
Initially developed for natural language processing (NLP), Transformer model is now widely used for speech processing tasks such as speaker recognition, due to its powerful sequenc…
EfficientTDNN: Efficient Architecture Search for Speaker Recognition
Rui Wang, Zhihua Wei, Haoran Duan +3
Convolutional neural networks (CNNs), such as the time-delay neural network (TDNN), have shown their remarkable capability in learning speaker embedding. However, they meanwhile br…
Tongji University Team for the VoxCeleb Speaker Recognition Challenge 2020
Rui Wang, Zhihua Wei, Yibin Zhan +1
In this report, we describe the submission of Tongji University team to the CLOSE track of the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2020 at Interspeech 2020. We investig…