17 citations · 18 across the 7 of their papers we have counts for
4 papers · 1 filter
Graph Neural Network Backend for Speaker Recognition
Liang He, Ruida Li, Mengqi Niu
Currently, most speaker recognition backends, such as cosine, linear discriminant analysis (LDA), or probabilistic linear discriminant analysis (PLDA), make decisions by calculatin…
I4U System Description for NIST SRE'20 CTS Challenge
Kong Aik Lee, Tomi Kinnunen, Daniele Colibro +23
This manuscript describes the I4U submission to the 2020 NIST Speaker Recognition Evaluation (SRE'20) Conversational Telephone Speech (CTS) Challenge. The I4U's submission was resu…
Adaptive Multi-scale Detection of Acoustic Events
Wenhao Ding, Liang He
The goal of acoustic (or sound) events detection (AED or SED) is to predict the temporal position of target events in given audio segments. This task plays a significant role in sa…
Latent Class Model with Application to Speaker Diarization
Liang He, Xianhong Chen, Can Xu +3
In this paper, we apply a latent class model (LCM) to the task of speaker diarization. LCM is similar to Patrick Kenny's variational Bayes (VB) method in that it uses soft informat…