1 paper
Junhui Zhang, Qianhui Xu, Qingxiang Guo +4
While recent text-to-speech (TTS) models achieve high naturalness, controlling fine-grained expression via natural-language instructions remains challenging. We introduce Poly- Ins…