2 citations · 7 across the 20 of their papers we have counts for
4 papers · 1 filter
Voice Conversion for Lombard Speaking Style with Implicit and Explicit Acoustic Feature Conditioning
Dominika Woszczyk, Manuel Sam Ribeiro, Thomas Merritt +1
Text-to-Speech (TTS) systems in Lombard speaking style can improve the overall intelligibility of speech, useful for hearing loss and noisy conditions. However, training those mode…
AE-Flow: AutoEncoder Normalizing Flow
Jakub Mosiński, Piotr Biliński, Thomas Merritt +2
Recently normalizing flows have been gaining traction in text-to-speech (TTS) and voice conversion (VC) due to their state-of-the-art (SOTA) performance. Normalizing flows are unsu…
Creating New Voices using Normalizing Flows
Piotr Bilinski, Thomas Merritt, Abdelhamid Ezzerg +5
Creating realistic and natural-sounding synthetic speech remains a big challenge for voice identities unseen during training. As there is growing interest in synthesizing voices of…
Non-Autoregressive TTS with Explicit Duration Modelling for Low-Resource Highly Expressive Speech
Raahil Shah, Kamil Pokora, Abdelhamid Ezzerg +5
Whilst recent neural text-to-speech (TTS) approaches produce high-quality speech, they typically require a large amount of recordings from the target speaker. In previous work, a 3…