9 citations · 9 across the 3 of their papers we have counts for
4 papers
A study on the efficacy of model pre-training in developing neural text-to-speech system
Guangyan Zhang, Yichong Leng, Daxin Tan +5
In the development of neural text-to-speech systems, model pre-training with a large amount of non-target speakers' data is a common approach. However, in terms of ultimately achie…
Applying the Information Bottleneck Principle to Prosodic Representation Learning
Guangyan Zhang, Ying Qin, Daxin Tan +1
This paper describes a novel design of a neural network-based speech generation model for learning prosodic representation.The problem of representation learning is formulated acco…
AdaSpeech 3: Adaptive Text to Speech for Spontaneous Style
Yuzi Yan, Xu Tan, Bohan Li +6
While recent text to speech (TTS) models perform very well in synthesizing reading-style (e.g., audiobook) speech, it is still challenging to synthesize spontaneous-style speech (e…
CUHK-EE Voice Cloning System for ICASSP 2021 M2VoC Challenge
Daxin Tan, Hingpang Huang, Guangyan Zhang +1
This paper presents the CUHK-EE voice cloning system for ICASSP 2021 M2VoC challenge. The challenge provides two Mandarin speech corpora: the AIShell-3 corpus of 218 speakers with…