1 paper · 1 filter
Tu Vo, Sheir Zaheer, Chan Y. Park
Contrastive audio-language models such as CLAP enable zero-shot audio classification: a sound is labelled by matching its embedding to text prompt embeddings, with no labelled audi…