4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.SD2025
CrossMuSim: A Cross-Modal Framework for Music Similarity Retrieval with LLM-Powered Text Description Sourcing and Mining
Tristan Tsoi, Jiajun Deng, Yaolong Ju +3
Music similarity retrieval is fundamental for managing and exploring relevant content from large collections in streaming platforms. This paper presents a novel cross-modal contras…
cs.CL2023★ 4 cited
WikiMuTe: A web-sourced dataset of semantic descriptions for music audio
Benno Weck, Holger Kirchhoff, Peter Grosche +1
Multi-modal deep learning techniques for matching free-form text with music have shown promising results in the field of Music Information Retrieval (MIR). Prior work is often base…