1 paper · 1 filter
Seungheon Doh, Minhee Lee, Sangmoon Lee +2
We present VTMR, a two-stage framework for Video-To-Music Recommendation. In Stage~1, VTMR aligns comprehensive video and music signals in a joint audio-visual-text representation…