1 paper · 1 filter
Jinhyeok Yang, Hyeongju Kim, Yechan Yu +3
While flow-matching text-to-speech (TTS) achieves strong zero-shot speaker similarity and naturalness, it remains susceptible to content fidelity issues, particularly skip and repe…