1 citations · 1 across the 3 of their papers we have counts for
3 papers
Investigating self-supervised features for expressive, multilingual voice conversion
Álvaro Martín-Cortinas, Daniel Sáez-Trigueros, Grzegorz Beringer +7
Voice conversion (VC) systems are widely used for several applications, from speaker anonymisation to personalised speech synthesis. Supervised approaches learn a mapping between d…
Del Visual al Auditivo: Sonorización de Escenas Guiada por Imagen
María Sánchez, Laura Fernández, Julián Arias +6
Recent advances in image, video, text and audio generative techniques, and their use by the general public, are leading to new forms of content generation. Usually, each modality w…
Low-data? No problem: low-resource, language-agnostic conversational text-to-speech via F0-conditioned data augmentation
Giulia Comini, Goeric Huybrechts, Manuel Sam Ribeiro +2
The availability of data in expressive styles across languages is limited, and recording sessions are costly and time consuming. To overcome these issues, we demonstrate how to bui…