85 citations · 85 across the 3 of their papers we have counts for
9 papers
Overlap-aware low-latency online speaker diarization based on end-to-end local segmentation
Juan M. Coria, Hervé Bredin, Sahar Ghannay +1
We propose to address online speaker diarization as a combination of incremental clustering and local diarization applied to a rolling buffer updated every 500ms. Every single step…
End-to-end speaker segmentation for overlap-aware resegmentation
Hervé Bredin, Antoine Laurent
Speaker segmentation consists in partitioning a conversation between one or more speakers into speaker turns. Usually addressed as the late combination of three sub-tasks (voice ac…
A Comparison of Metric Learning Loss Functions for End-To-End Speaker Verification
Juan M. Coria, Hervé Bredin, Sahar Ghannay +1
Despite the growing popularity of metric learning approaches, very little work has attempted to perform a fair comparison of these techniques for speaker verification. We try to fi…
Speaker detection in the wild: Lessons learned from JSALT 2019
Paola Garcia, Jesus Villalba, Herve Bredin +21
This paper presents the problems and solutions addressed at the JSALT workshop when using a single microphone for speaker detection in adverse scenarios. The main focus was to tack…
The Speed Submission to DIHARD II: Contributions & Lessons Learned
Md Sahidullah, Jose Patino, Samuele Cornell +11
This paper describes the speaker diarization systems developed for the Second DIHARD Speech Diarization Challenge (DIHARD II) by the Speed team. Besides describing the system, whic…
pyannote.audio: neural building blocks for speaker diarization
Hervé Bredin, Ruiqing Yin, Juan Manuel Coria +7
We introduce pyannote.audio, an open-source toolkit written in Python for speaker diarization. Based on PyTorch machine learning framework, it provides a set of trainable end-to-en…