1 paper
John Janiczek, Dading Chong, Dongyang Dai +4
A text-to-speech (TTS) model trained to reconstruct speech given text tends towards predictions that are close to the average characteristics of a dataset, failing to model the var…