2 papers
cs.SD2026
IndexTTS 2.5 Technical Report
Yunpei Li, Xun Zhou, Jinchao Wang +11
In prior work, we introduced IndexTTS 2, a zero-shot neural text-to-speech foundation model comprising two core components: a transformer-based Text-to-Semantic (T2S) module and a…
cs.SD2025
Index-ASR Technical Report
Zheshu Song, Lu Wang, Wei Deng +3
Automatic speech recognition (ASR) has witnessed remarkable progress in recent years, largely driven by the emergence of LLM-based ASR paradigm. Despite their strong performance on…