303 citations · 631 across the 26 of their papers we have counts for
4 papers · 1 filter
Target-speaker Voice Activity Detection with Improved I-Vector Estimation for Unknown Number of Speaker
Maokui He, Desh Raj, Zili Huang +3
Target-speaker voice activity detection (TS-VAD) has recently shown promising results for speaker diarization on highly overlapped speech. However, the original model requires a fi…
Separation Guided Speaker Diarization in Realistic Mismatched Conditions
Shu-Tong Niu, Jun Du, Lei Sun +1
We propose a separation guided speaker diarization (SGSD) approach by fully utilizing a complementarity of speech separation and speaker clustering. Since the conventional clusteri…
AISHELL-4: An Open Source Dataset for Speech Enhancement, Separation, Recognition and Speaker Diarization in Conference Scenario
Yihui Fu, Luyao Cheng, Shubo Lv +10
In this paper, we present AISHELL-4, a sizable real-recorded Mandarin speech dataset collected by 8-channel circular microphone array for speech processing in conference scenario.…
USTC-NELSLIP System Description for DIHARD-III Challenge
Yuxuan Wang, Maokui He, Shutong Niu +6
This system description describes our submission system to the Third DIHARD Speech Diarization Challenge. Besides the traditional clustering based system, the innovation of our sys…