8 citations · 12 across the 3 of their papers we have counts for
3 papers
eess.AS2019★ 3 cited
Training Multi-Speaker Neural Text-to-Speech Systems using Speaker-Imbalanced Speech Corpora
Hieu-Thi Luong, Xin Wang, Junichi Yamagishi +1
When the available data of a target speaker is insufficient to train a high quality speaker-dependent neural text-to-speech (TTS) system, we can combine data from multiple speakers…
eess.AS2019★ 8 cited
Joint training framework for text-to-speech and voice conversion using multi-source Tacotron and WaveNet
Mingyang Zhang, Xin Wang, Fuming Fang +2
We investigated the training of a shared model for both text-to-speech (TTS) and voice conversion (VC) tasks. We propose using an extended model architecture of Tacotron, that is a…
cs.LG2018★ 1 cited
Conditional Graph Neural Processes: A Functional Autoencoder Approach
Marcel Nassar, Xin Wang, Evren Tumer
We introduce a novel encoder-decoder architecture to embed functional processes into latent vector spaces. This embedding can then be decoded to sample the encoded functions over a…