35 citations · 64 across the 3 of their papers we have counts for
3 papers
eess.AS2023★ 16 cited
Mega-TTS: Zero-Shot Text-to-Speech at Scale with Intrinsic Inductive Bias
Ziyue Jiang, Yi Ren, Zhenhui Ye +9
Scaling text-to-speech to a large and wild dataset has been proven to be highly effective in achieving timbre and speech style generalization, particularly in zero-shot TTS. Howeve…
cs.SD2022★ 13 cited
ReLyMe: Improving Lyric-to-Melody Generation by Incorporating Lyric-Melody Relationships
Chen Zhang, Luchin Chang, Songruoyao Wu +4
Lyric-to-melody generation, which generates melody according to given lyrics, is one of the most important automatic music composition tasks. With the rapid development of deep lea…
eess.AS2022★ 35 cited
NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality
Xu Tan, Jiawei Chen, Haohe Liu +11
Text to speech (TTS) has made rapid progress in both academia and industry in recent years. Some questions naturally arise that whether a TTS system can achieve human-level quality…