5 citations · 10 across the 4 of their papers we have counts for
4 papers
Building Bilingual and Code-Switched Voice Conversion with Limited Training Data Using Embedding Consistency Loss
Yaogen Yang, Haozhe Zhang, Xiaoyi Qin +4
Building cross-lingual voice conversion (VC) systems for multiple speakers and multiple languages has been a challenging task for a long time. This paper describes a parallel non-a…
Exploring Voice Conversion based Data Augmentation in Text-Dependent Speaker Verification
Xiaoyi Qin, Yaogen Yang, Lin Yang +3
In this paper, we focus on improving the performance of the text-dependent speaker verification system in the scenario of limited training data. The speaker verification system dee…
Cross-lingual Multispeaker Text-to-Speech under Limited-Data Scenario
Zexin Cai, Yaogen Yang, Ming Li
Modeling voices for multiple speakers and multiple languages in one text-to-speech system has been a challenge for a long time. This paper presents an extension on Tacotron2 to ach…
Polyphone Disambiguation for Mandarin Chinese Using Conditional Neural Network with Multi-level Embedding Features
Zexin Cai, Yaogen Yang, Chuxiong Zhang +2
This paper describes a conditional neural network architecture for Mandarin Chinese polyphone disambiguation. The system is composed of a bidirectional recurrent neural network com…