1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Siyi Zhou, Yiquan Zhou, Yi He +4
Existing autoregressive large-scale text-to-speech (TTS) models have advantages in speech naturalness, but their token-by-token generation mechanism makes it difficult to precisely…