35 citations · 38 across the 3 of their papers we have counts for
3 papers
eess.AS2022★ 35 cited
NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality
Xu Tan, Jiawei Chen, Haohe Liu +11
Text to speech (TTS) has made rapid progress in both academia and industry in recent years. Some questions naturally arise that whether a TTS system can achieve human-level quality…
eess.AS2021★ 3 cited
PDAugment: Data Augmentation by Pitch and Duration Adjustments for Automatic Lyrics Transcription
Chen Zhang, Jiaxing Yu, LuChin Chang +4
Automatic lyrics transcription (ALT), which can be regarded as automatic speech recognition (ASR) on singing voice, is an interesting and practical topic in academia and industry.…
cs.CL2021
A General Framework for Learning Prosodic-Enhanced Representation of Rap Lyrics
Hongru Liang, Haozheng Wang, Qian Li +5
Learning and analyzing rap lyrics is a significant basis for many web applications, such as music recommendation, automatic music categorization, and music information retrieval, d…