708 citations · 1.5k across the 27 of their papers we have counts for
1 paper · 2 filters
Sercan Arik, Gregory Diamos, Andrew Gibiansky +5
We introduce a technique for augmenting neural text-to-speech (TTS) with lowdimensional trainable speaker embeddings to generate different voices from a single model. As a starting…