7 citations · 13 across the 4 of their papers we have counts for
1 paper · 1 filter
Benno Weck, Miguel Pérez Fernández, Holger Kirchhoff +1
We present an analysis of large-scale pretrained deep learning models used for cross-modal (text-to-audio) retrieval. We use embeddings extracted by these models in a metric learni…