9 citations · 9 across the 3 of their papers we have counts for
3 papers
cs.SD2021
Full Attention Bidirectional Deep Learning Structure for Single Channel Speech Enhancement
Yuzi Yan, Wei-Qiang Zhang, Michael T. Johnson
As the cornerstone of other important technologies, such as speech recognition and speech synthesis, speech enhancement is a critical area in audio signal processing. In this paper…
cs.SD2021★ 9 cited
AdaSpeech 3: Adaptive Text to Speech for Spontaneous Style
Yuzi Yan, Xu Tan, Bohan Li +6
While recent text to speech (TTS) models perform very well in synthesizing reading-style (e.g., audiobook) speech, it is still challenging to synthesize spontaneous-style speech (e…
cs.SD2021
AdaSpeech 2: Adaptive Text to Speech with Untranscribed Data
Yuzi Yan, Xu Tan, Bohan Li +4
Text to speech (TTS) is widely used to synthesize personal voice for a target speaker, where a well-trained source TTS model is fine-tuned with few paired adaptation data (speech a…