15 citations · 15 across the 2 of their papers we have counts for
3 papers
eess.AS2021
Controllable Context-aware Conversational Speech Synthesis
Jian Cong, Shan Yang, Na Hu +3
In spoken conversations, spontaneous behaviors like filled pause and prolongations always happen. Conversational partner tends to align features of their speech with their interloc…
cs.SD2021★ 15 cited
VARA-TTS: Non-Autoregressive Text-to-Speech Synthesis based on Very Deep VAE with Residual Attention
Peng Liu, Yuewen Cao, Songxiang Liu +4
This paper proposes VARA-TTS, a non-autoregressive (non-AR) text-to-speech (TTS) model using a very deep Variational Autoencoder (VDVAE) with Residual Attention mechanism, which re…
eess.AS2019
Maximizing Mutual Information for Tacotron
Peng Liu, Xixin Wu, Shiyin Kang +3
End-to-end speech synthesis methods already achieve close-to-human quality performance. However compared to HMM-based and NN-based frame-to-frame regression methods, they are prone…