183 citations · 507 across the 18 of their papers we have counts for
8 papers · 1 filter
Can we use Common Voice to train a Multi-Speaker TTS system?
Sewade Ogun, Vincent Colotte, Emmanuel Vincent
Training of multi-speaker text-to-speech (TTS) systems relies on curated datasets based on high-quality recordings or audiobooks. Such datasets often lack speaker diversity and are…
UIAI System for Short-Duration Speaker Verification Challenge 2020
Md Sahidullah, Achintya Kumar Sarkar, Ville Vestman +5
In this work, we present the system description of the UIAI entry for the short-duration speaker verification (SdSV) challenge 2020. Our focus is on Task 1 dedicated to text-depend…
LibriMix: An Open-Source Dataset for Generalizable Speech Separation
Joris Cosentino, Manuel Pariente, Samuele Cornell +2
In recent years, wsj0-2mix has become the reference dataset for single-channel speech separation. Most deep learning-based speech separation models today are benchmarked on it. How…
Design Choices for X-vector Based Speaker Anonymization
Brij Mohan Lal Srivastava, Natalia Tomashenko, Xin Wang +5
The recently proposed x-vector based anonymization scheme converts any input voice into that of a random pseudo-speaker. In this paper, we present a flexible pseudo-speaker selecti…
Asteroid: the PyTorch-based audio source separation toolkit for researchers
Manuel Pariente, Samuele Cornell, Joris Cosentino +11
This paper describes Asteroid, the PyTorch-based audio source separation toolkit for researchers. Inspired by the most successful neural source separation systems, it provides all…
Foreground-Background Ambient Sound Scene Separation
Michel Olvera, Emmanuel Vincent, Romain Serizel +1
Ambient sound scenes typically comprise multiple short events occurring on top of a somewhat stationary background. We consider the task of separating these events from the backgro…