35 citations · 40 across the 3 of their papers we have counts for
3 papers
eess.AS2022★ 35 cited
NaturalSpeech: End-to-End Text to Speech Synthesis with Human-Level Quality
Xu Tan, Jiawei Chen, Haohe Liu +11
Text to speech (TTS) has made rapid progress in both academia and industry in recent years. Some questions naturally arise that whether a TTS system can achieve human-level quality…
eess.AS2020★ 3 cited
s-Transformer: Segment-Transformer for Robust Neural Speech Synthesis
Xi Wang, Huaiping Ming, Lei He +1
Neural end-to-end text-to-speech (TTS) , which adopts either a recurrent model, e.g. Tacotron, or an attention one, e.g. Transformer, to characterize a speech utterance, has achiev…
eess.AS2019★ 2 cited
Forward-Backward Decoding for Regularizing End-to-End TTS
Yibin Zheng, Xi Wang, Lei He +4
Neural end-to-end TTS can generate very high-quality synthesized speech, and even close to human recording within similar domain text. However, it performs unsatisfactory when scal…