7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.SD2024
Do Captioning Metrics Reflect Music Semantic Alignment?
Jinwoo Lee, Kyogu Lee
Music captioning has emerged as a promising task, fueled by the advent of advanced language generation models. However, the evaluation of music captioning relies heavily on traditi…
cs.SD2022★ 7 cited
Exploring Train and Test-Time Augmentations for Audio-Language Learning
Eungbeom Kim, Jinhee Kim, Yoori Oh +5
In this paper, we aim to unveil the impact of data augmentation in audio-language multi-modal learning, which has not been explored despite its importance. We explore various augme…