1 paper
Xiaoxue Gao, Huayun Zhang, Nancy F. Chen
Existing expressive text-to-speech (TTS) systems primarily model a limited set of categorical emotions, whereas human conversations extend far beyond these predefined emotions, mak…