1 paper
Yunpei Li, Xun Zhou, Jinchao Wang +11
In prior work, we introduced IndexTTS 2, a zero-shot neural text-to-speech foundation model comprising two core components: a transformer-based Text-to-Semantic (T2S) module and a…