activity
20192021
most citedSpEx: Multi-Scale Time Domain Speaker Extraction Network

158 citations · 227 across the 11 of their papers we have counts for

collaborators

12 papers

eess.AS20213 cited

Improving Channel Decorrelation for Multi-Channel Target Speech Extraction

Jiangyu Han, Wei Rao, Yannan Wang +1

Target speech extraction has attracted widespread attention. When microphone arrays are available, the additional spatial information can be helpful in extracting the target speech…

eess.AS20217 cited

INTERSPEECH 2021 ConferencingSpeech Challenge: Towards Far-field Multi-Channel Speech Enhancement for Video Conferencing

Wei Rao, Yihui Fu, Yanxin Hu +11

The ConferencingSpeech 2021 challenge is proposed to stimulate research on far-field multi-channel speech enhancement for video conferencing. The challenge consists of two separate…

eess.AS20212 cited

Target Speaker Verification with Selective Auditory Attention for Single and Multi-talker Speech

Chenglin Xu, Wei Rao, Jibin Wu +1

Speaker verification has been studied mostly under the single-talker condition. It is adversely affected in the presence of interference speakers. Inspired by the study on target s…

cs.SD20201 cited

Adversarial Training for Multi-domain Speaker Recognition

Qing Wang, Wei Rao, Pengcheng Guo +1

In real-life applications, the performance of speaker recognition systems always degrades when there is a mismatch between training and evaluation data. Many domain adaptation meth…

eess.AS2020

HLT-NUS Submission for NIST 2019 Multimedia Speaker Recognition Evaluation

Rohan Kumar Das, Ruijie Tao, Jichen Yang +3

This work describes the speaker verification system developed by Human Language Technology Laboratory, National University of Singapore (HLT-NUS) for 2019 NIST Multimedia Speaker R…

eess.AS20203 cited

The INTERSPEECH 2020 Far-Field Speaker Verification Challenge

Xiaoyi Qin, Ming Li, Hui Bu +4

The INTERSPEECH 2020 Far-Field Speaker Verification Challenge (FFSVC 2020) addresses three different research problems under well-defined conditions: far-field text-dependent speak…