1 paper
Chenpeng Du, Yiwei Guo, Xie Chen +1
The mainstream neural text-to-speech(TTS) pipeline is a cascade system, including an acoustic model(AM) that predicts acoustic feature from the input transcript and a vocoder that…