1 paper
Heyang Xue, Xuchen Song, Yu Tang +4
Description-based text-to-speech (TTS) models exhibit strong performance on in-domain text descriptions, i.e., those encountered during training. However, in real-world application…