21 citations · 28 across the 2 of their papers we have counts for
1 paper · 1 filter
Guangzhi Sun, Yu Zhang, Ron J. Weiss +3
This paper proposes a hierarchical, fine-grained and interpretable latent variable model for prosody based on the Tacotron 2 text-to-speech model. It achieves multi-resolution mode…