1 paper · 1 filter
Peijie Chen, Wenhao Guan, Weijie Wu +7
Zero-shot text-to-speech (TTS) relies on robust speech representations. However, current speech tokenizers face a fundamental trade-off: acoustic codecs preserve high-fidelity audi…