VoxSRC 2021: The Third VoxCeleb Speaker Recognition Challenge
arXiv:2201.04583
Abstract
The third instalment of the VoxCeleb Speaker Recognition Challenge was held in conjunction with Interspeech 2021. The aim of this challenge was to assess how well current speaker recognition technology is able to diarise and recognise speakers in unconstrained or `in the wild' data. The challenge consisted of: (i) the provision of publicly available speaker recognition and diarisation data from YouTube videos together with ground truth annotation and standardised evaluation software; and (ii) a virtual public challenge and workshop held at Interspeech 2021. This paper outlines the challenge, and describes the baselines, methods and results. We conclude with a discussion on the new multi-lingual focus of VoxSRC 2021, and on the progression of the challenge since the previous two editions.
arXiv admin note: substantial text overlap with arXiv:2012.06867
Cited by in corpus (5)
- The VoxCeleb Speaker Recognition Challenge: A Retrospective
- TSUP Speaker Diarization System for Conversational Short-phrase Speaker Diarization Challenge
- Deep Neural Networks for Automatic Speaker Recognition Do Not Learn Supra-Segmental Temporal Features
- A Teacher-Student approach for extracting informative speaker embeddings from speech mixtures
- HeightCeleb - an enrichment of VoxCeleb dataset with speaker height information