32 citations · 68 across the 7 of their papers we have counts for
Showing 2021Show all
2 papers · 1 filter
eess.AS2021
Multi-Scale Spectrogram Modelling for Neural Text-to-Speech
Ammar Abbas, Bajibabu Bollepalli, Alexis Moinet +6
We propose a novel Multi-Scale Spectrogram (MSS) modelling approach to synthesise speech with an improved coarse and fine-grained prosody. We present a generic multi-scale spectrog…
eess.AS2021
A learned conditional prior for the VAE acoustic space of a TTS system
Penny Karanasou, Sri Karlapati, Alexis Moinet +5
Many factors influence speech yielding different renditions of a given sentence. Generative models, such as variational autoencoders (VAEs), capture this variability and allow mult…