1 citations · 1 across the 1 of their papers we have counts for
1 paper
Benno Weck, Miguel Pérez Fernández, Holger Kirchhoff +1
We present an analysis of large-scale pretrained deep learning models used for cross-modal (text-to-audio) retrieval. We use embeddings extracted by these models in a metric learni…