activity
20202022
most citedSpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing

30 citations · 55 across the 4 of their papers we have counts for

collaborators

5 papers

eess.AS2022

LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT

Rui Wang, Qibing Bai, Junyi Ao +6

Self-supervised speech representation learning has shown promising results in various speech processing tasks. However, the pre-trained models, e.g., HuBERT, are storage-intensive…

eess.AS2021★ 30 cited

SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing

Junyi Ao, Rui Wang, Long Zhou +11

Motivated by the success of T5 (Text-To-Text Transfer Transformer) in pre-trained natural language processing models, we propose a unified-modal SpeechT5 framework that explores th…

eess.AS2021★ 3 cited

Multi-View Self-Attention Based Transformer for Speaker Recognition

Rui Wang, Junyi Ao, Long Zhou +5

Initially developed for natural language processing (NLP), Transformer model is now widely used for speech processing tasks such as speaker recognition, due to its powerful sequenc…

eess.AS2021★ 22 cited

EfficientTDNN: Efficient Architecture Search for Speaker Recognition

Rui Wang, Zhihua Wei, Haoran Duan +3

Convolutional neural networks (CNNs), such as the time-delay neural network (TDNN), have shown their remarkable capability in learning speaker embedding. However, they meanwhile br…

eess.AS2020

Tongji University Team for the VoxCeleb Speaker Recognition Challenge 2020

Rui Wang, Zhihua Wei, Yibin Zhan +1

In this report, we describe the submission of Tongji University team to the CLOSE track of the VoxCeleb Speaker Recognition Challenge (VoxSRC) 2020 at Interspeech 2020. We investig…