11 citations · 28 across the 8 of their papers we have counts for
3 papers · 1 filter
The JHU submission to VoxSRC-21: Track 3
Jejin Cho, Jesus Villalba, Najim Dehak
This technical report describes Johns Hopkins University speaker recognition system submitted to Voxceleb Speaker Recognition Challenge 2021 Track 3: Self-supervised speaker verifi…
Learning Speaker Embedding from Text-to-Speech
Jaejin Cho, Piotr Zelasko, Jesus Villalba +2
Zero-shot multi-speaker Text-to-Speech (TTS) generates target speaker voices given an input text and the corresponding speaker embedding. In this work, we investigate the effective…
Deep neural networks for emotion recognition combining audio and transcripts
Jaejin Cho, Raghavendra Pappagari, Purva Kulkarni +3
In this paper, we propose to improve emotion recognition by combining acoustic information and conversation transcripts. On the one hand, an LSTM network was used to detect emotion…