activity
20192022
most citedTraditional Machine Learning for Pitch Detection

32 citations · 66 across the 5 of their papers we have counts for

collaborators

7 papers

eess.AS2022

Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module

Adam Gabryś, Goeric Huybrechts, Manuel Sam Ribeiro +6

State-of-the-art text-to-speech (TTS) systems require several hours of recorded speech data to generate high-quality synthetic speech. When using reduced amounts of training data,…

eess.AS20222 cited

Cross-speaker style transfer for text-to-speech using data augmentation

Manuel Sam Ribeiro, Julian Roth, Giulia Comini +3

We address the problem of cross-speaker style transfer for text-to-speech (TTS) using data augmentation via voice conversion. We assume to have a corpus of neutral non-expressive d…

cs.SD20211 cited

Non-Autoregressive TTS with Explicit Duration Modelling for Low-Resource Highly Expressive Speech

Raahil Shah, Kamil Pokora, Abdelhamid Ezzerg +5

Whilst recent neural text-to-speech (TTS) approaches produce high-quality speech, they typically require a large amount of recordings from the target speaker. In previous work, a 3…

eess.AS2021

EmoCat: Language-agnostic Emotional Voice Conversion

Bastian Schnell, Goeric Huybrechts, Bartek Perz +2

Emotional voice conversion models adapt the emotion in speech without changing the speaker identity or linguistic content. They are less data hungry than text-to-speech models and…

eess.AS2020

Low-resource expressive text-to-speech using data augmentation

Goeric Huybrechts, Thomas Merritt, Giulia Comini +3

While recent neural text-to-speech (TTS) systems perform remarkably well, they typically require a substantial amount of recordings from the target speaker reading in the desired s…

cs.SD202031 cited

Voice Conversion for Whispered Speech Synthesis

Marius Cotescu, Thomas Drugman, Goeric Huybrechts +2

We present an approach to synthesize whisper by applying a handcrafted signal processing recipe and Voice Conversion (VC) techniques to convert normally phonated speech to whispere…