20 citations · 22 across the 3 of their papers we have counts for
3 papers
eess.AS2021★ 2 cited
Target-speaker Voice Activity Detection with Improved I-Vector Estimation for Unknown Number of Speaker
Maokui He, Desh Raj, Zili Huang +3
Target-speaker voice activity detection (TS-VAD) has recently shown promising results for speaker diarization on highly overlapped speech. However, the original model requires a fi…
cs.SD2021★ 20 cited
USTC-NELSLIP System Description for DIHARD-III Challenge
Yuxuan Wang, Maokui He, Shutong Niu +6
This system description describes our submission system to the Third DIHARD Speech Diarization Challenge. Besides the traditional clustering based system, the innovation of our sys…
eess.AS2020
Integration of speech separation, diarization, and recognition for multi-speaker meetings: System description, comparison, and analysis
Desh Raj, Pavel Denisov, Zhuo Chen +11
Multi-speaker speech recognition of unsegmented recordings has diverse applications such as meeting transcription and automatic subtitle generation. With technical advances in syst…