1 citations · 1 across the 1 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
Ludovic Tuncay, Etienne Labbé, Emmanouil Benetos +1
Building on the Joint-Embedding Predictive Architecture (JEPA) paradigm, a recent self-supervised learning framework that predicts latent representations of masked regions in high-…
cs.SD2022★ 1 cited
Is my automatic audio captioning system so bad? spider-max: a metric to consider several caption candidates
Etienne Labbé, Thomas Pellegrini, Julien Pinquier
Automatic Audio Captioning (AAC) is the task that aims to describe an audio signal using natural language. AAC systems take as input an audio signal and output a free-form text sen…