VoxSRC 2020: The Second VoxCeleb Speaker Recognition Challenge
arXiv:2012.06867
Abstract
We held the second installment of the VoxCeleb Speaker Recognition Challenge in conjunction with Interspeech 2020. The goal of this challenge was to assess how well current speaker recognition technology is able to diarise and recognize speakers in unconstrained or `in the wild' data. It consisted of: (i) a publicly available speaker recognition and diarisation dataset from YouTube videos together with ground truth annotation and standardised evaluation software; and (ii) a virtual public challenge and workshop held at Interspeech 2020. This paper outlines the challenge, and describes the baselines, methods used, and results. We conclude with a discussion of the progress over the first installment of the challenge.
References in corpus (9)
- Conformer: Convolution-augmented Transformer for Speech Recognition
- CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings
- BUT System Description to VoxCeleb Speaker Recognition Challenge 2019
- A Framework For Contrastive Self-Supervised Learning And Designing A New Approach
- VoxSRC 2019: The first VoxCeleb Speaker Recognition Challenge
- The VOiCES from a Distance Challenge 2019 Evaluation Plan
- Semi-Supervised Contrastive Learning with Generalized Contrastive Loss and Its Application to Speaker Recognition
- The xx205 System for the VoxCeleb Speaker Recognition Challenge 2020
- The UPC Speaker Verification System Submitted to VoxCeleb Speaker Recognition Challenge 2020 (VoxSRC-20)