297 citations · 666 across the 18 of their papers we have counts for
4 papers · 1 filter
Data Summarization via Bilevel Optimization
Zalán Borsos, Mojmír Mutný, Marco Tagliasacchi +1
The increasing availability of massive data sets poses a series of challenges for machine learning. Prominent among these is the need to learn models under hardware or human resour…
SoundStream: An End-to-End Neural Audio Codec
Neil Zeghidour, Alejandro Luebs, Ahmed Omran +2
We present SoundStream, a novel neural audio codec that can efficiently compress speech, music and general audio at bitrates normally targeted by speech-tailored codecs. SoundStrea…
Self-Supervised Learning from Automatically Separated Sound Scenes
Eduardo Fonseca, Aren Jansen, Daniel P. W. Ellis +7
Real-world sound scenes consist of time-varying collections of sound sources, each generating characteristic sound events that are mixed together in audio recordings. The associati…
LEAF: A Learnable Frontend for Audio Classification
Neil Zeghidour, Olivier Teboul, Félix de Chaumont Quitry +1
Mel-filterbanks are fixed, engineered audio features which emulate human perception and have been used through the history of audio understanding up to today. However, their undeni…