3 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.AI2023★ 1 cited
Rhythm-controllable Attention with High Robustness for Long Sentence Speech Synthesis
Dengfeng Ke, Yayue Deng, Yukang Jia +6
Regressive Text-to-Speech (TTS) system utilizes attention mechanism to generate alignment between text and acoustic feature sequence. Alignment determines synthesis robustness (e.g…
cs.SD2023★ 3 cited
M2-CTTS: End-to-End Multi-scale Multi-modal Conversational Text-to-Speech Synthesis
Jinlong Xue, Yayue Deng, Fengping Wang +5
Conversational text-to-speech (TTS) aims to synthesize speech with proper prosody of reply based on the historical conversation. However, it is still a challenge to comprehensively…