1 paper
Seongyeon Park, Myungseo Song, Bohyung Kim +1
Neural text-to-speech (TTS) models can synthesize natural human speech when trained on large amounts of transcribed speech. However, collecting such large-scale transcribed data is…