1 paper
Ziyang Zhang, Yifan Gao, Xuenan Xu +3
Text-to-Speech (TTS) is inherently a "one-to-many" mapping characterized by intrinsic uncertainty, yet current paradigms often oversimplify it into a deterministic regression task.…