1 paper
Biel Tura Vecino, Adam GabryÅ, Daniel MÄ twicki +4
Recent works have shown that modelling raw waveform directly from text in an end-to-end (E2E) fashion produces more natural-sounding speech than traditional neural text-to-speech (…