5 citations · 5 across the 1 of their papers we have counts for
4 papers
RW-Resnet: A Novel Speech Anti-Spoofing Model Using Raw Waveform
Youxuan Ma, Zongze Ren, Shugong Xu
In recent years, synthetic speech generated by advanced text-to-speech (TTS) and voice conversion (VC) systems has caused great harms to automatic speaker verification (ASV) system…
A Study on Angular Based Embedding Learning for Text-independent Speaker Verification
Zhiyong Chen, Zongze Ren, Shugong Xu
Learning a good speaker embedding is important for many automatic speaker recognition tasks, including verification, identification and diarization. The embeddings learned by softm…
Two-stage Training for Chinese Dialect Recognition
Zongze Ren, Guofu Yang, Shugong Xu
In this paper, we present a two-stage language identification (LID) system based on a shallow ResNet14 followed by a simple 2-layer recurrent neural network (RNN) architecture, whi…
Triplet Based Embedding Distance and Similarity Learning for Text-independent Speaker Verification
Zongze Ren, Zhiyong Chen, Shugong Xu
Speaker embeddings become growing popular in the text-independent speaker verification task. In this paper, we propose two improvements during the training stage. The improvements…