1 paper
Qingliang Meng, Yuqing Deng, Wei Liang +3
Current non-autoregressive (NAR) text-to-speech (TTS) systems still struggle to model diverse and speaker-dependent duration variation. We further observe that richer duration vari…