1 paper · 1 filter
Jiacheng Shi, Hongfei Du, Yangfan He +2
Emotional text-to-speech seeks to convey affect while preserving intelligibility and prosody, yet existing methods rely on coarse labels or proxy classifiers and receive only utter…