1 paper · 1 filter
Ali Asaria, Tony Salomone, Deep Gandhi
Open autoregressive neural-codec text-to-speech (TTS) models sound excellent on typical inputs yet suffer stochastic catastrophic failures: on a meaningful fraction of utterances t…