activity
20162022
most citedData Augmenting Contrastive Learning of Speech Representations in the Time Domain

93 citations · 190 across the 13 of their papers we have counts for

collaborators
Showing 2018Show all

5 papers · 1 filter

cs.CL2018

End-to-End Speech Recognition From the Raw Waveform

Neil Zeghidour, Nicolas Usunier, Gabriel Synnaeve +2

State-of-the-art speech recognition systems rely on fixed, hand-crafted features such as mel-filterbanks to preprocess the waveform before the training pipeline. In this paper, we…

cs.CL2018

Sampling strategies in Siamese Networks for unsupervised speech representation learning

Rachid Riad, Corentin Dancette, Julien Karadayi +3

Recent studies have investigated siamese network architectures for learning invariant speech representations using same-different side information at the word level. Here we invest…

cs.AI2018

IntPhys: A Framework and Benchmark for Visual Intuitive Physics Reasoning

Ronan Riochet, Mario Ynocente Castro, Mathieu Bernard +4

In order to reach human performance on complexvisual tasks, artificial systems need to incorporate a sig-nificant amount of understanding of the world in termsof macroscopic object…

cs.CL2018

Bayesian Models for Unit Discovery on a Very Low Resource Language

Lucas Ondel, Pierre Godard, Laurent Besacier +7

Developing speech technologies for low-resource languages has become a very active research field over the last decade. Among others, Bayesian models have shown some promising resu…

cs.CL2018

Linguistic unit discovery from multi-modal inputs in unwritten languages: Summary of the "Speaking Rosetta" JSALT 2017 Workshop

Odette Scharenborg, Laurent Besacier, Alan Black +16

We summarize the accomplishments of a multi-disciplinary workshop exploring the computational and scientific issues surrounding the discovery of linguistic units (subwords and word…