4 papers
The Speed Submission to DIHARD II: Contributions & Lessons Learned
Md Sahidullah, Jose Patino, Samuele Cornell +11
This paper describes the speaker diarization systems developed for the Second DIHARD Speech Diarization Challenge (DIHARD II) by the Speed team. Besides describing the system, whic…
pyannote.audio: neural building blocks for speaker diarization
Hervé Bredin, Ruiqing Yin, Juan Manuel Coria +7
We introduce pyannote.audio, an open-source toolkit written in Python for speaker diarization. Based on PyTorch machine learning framework, it provides a set of trainable end-to-en…
Cross-Lingual Contextual Word Embeddings Mapping With Multi-Sense Words In Mind
Zheng Zhang, Ruiqing Yin, Jun Zhu +1
Recent work in cross-lingual contextual word embedding learning cannot handle multi-sense words well. In this work, we explore the characteristics of contextual word embeddings and…
LSTM based Similarity Measurement with Spectral Clustering for Speaker Diarization
Qingjian Lin, Ruiqing Yin, Ming Li +2
More and more neural network approaches have achieved considerable improvement upon submodules of speaker diarization system, including speaker change detection and segment-wise sp…