1 paper
Zhichao Wu, Qiulin Li, Sixing Liu +1
In the Text-to-speech(TTS) task, the latent diffusion model has excellent fidelity and generalization, but its expensive resource consumption and slow inference speed have always b…