5 papers
Learning spectro-temporal representations of complex sounds with parameterized neural networks
Rachid Riad, Julien Karadayi, Anne-Catherine Bachoud-Lévi +1
Deep Learning models have become potential candidates for auditory neuroscience research, thanks to their recent successes on a variety of auditory tasks. Yet, these models often l…
The Zero Resource Speech Challenge 2020: Discovering discrete subword and word units
Ewan Dunbar, Julien Karadayi, Mathieu Bernard +6
We present the Zero Resource Speech Challenge 2020, which aims at learning speech representations from raw audio signals without any labels. It combines the data sets and metrics f…
Libri-Light: A Benchmark for ASR with Limited or No Supervision
Jacob Kahn, Morgane Rivière, Weiyi Zheng +12
We introduce a new collection of spoken English audio suitable for training speech recognition systems under limited or no supervision. It is derived from open-source audio books f…
The Zero Resource Speech Challenge 2019: TTS without T
Ewan Dunbar, Robin Algayres, Julien Karadayi +10
We present the Zero Resource Speech Challenge 2019, which proposes to build a speech synthesizer without any text or phonetic labels: hence, TTS without T (text-to-speech without t…
Sampling strategies in Siamese Networks for unsupervised speech representation learning
Rachid Riad, Corentin Dancette, Julien Karadayi +3
Recent studies have investigated siamese network architectures for learning invariant speech representations using same-different side information at the word level. Here we invest…