97 citations · 177 across the 6 of their papers we have counts for
6 papers
Clova Baseline System for the VoxCeleb Speaker Recognition Challenge 2020
Hee Soo Heo, Bong-Jin Lee, Jaesung Huh +1
This report describes our submission to the VoxCeleb Speaker Recognition Challenge (VoxSRC) at Interspeech 2020. We perform a careful analysis of speaker recognition models based o…
VoxSRC 2019: The first VoxCeleb Speaker Recognition Challenge
Joon Son Chung, Arsha Nagrani, Ernesto Coto +4
The VoxCeleb Speaker Recognition Challenge 2019 aimed to assess how well current speaker recognition technology is able to identify speakers in unconstrained or `in the wild' data.…
The sound of my voice: speaker representation loss for target voice separation
Seongkyu Mun, Soyeon Choe, Jaesung Huh +1
Content and style representations have been widely studied in the field of style transfer. In this paper, we propose a new loss function using speaker content representation for au…
Delving into VoxCeleb: environment invariant speaker recognition
Joon Son Chung, Jaesung Huh, Seongkyu Mun
Research in speaker recognition has recently seen significant progress due to the application of neural network models and the availability of new large-scale datasets. There has b…
Utterance-level Aggregation For Speaker Recognition In The Wild
Weidi Xie, Arsha Nagrani, Joon Son Chung +1
The objective of this paper is speaker recognition "in the wild"-where utterances may be of variable length and also contain irrelevant signals. Crucial elements in the design of d…
Signs in time: Encoding human motion as a temporal image
Joon Son Chung, Andrew Zisserman
The goal of this work is to recognise and localise short temporal signals in image time series, where strong supervision is not available for training. To this end we propose an im…