1 paper
Xiaofei Wang, Sefik Emre Eskimez, Manthan Thakker +8
Recently, zero-shot text-to-speech (TTS) systems, capable of synthesizing any speaker's voice from a short audio prompt, have made rapid advancements. However, the quality of the g…