1 paper
Benno Weck, Pablo Puentes, Andrea Poltronieri +2
The evaluation of music understanding in Large Audio-Language Models (LALMs) requires a rigorously defined benchmark that truly tests whether models can perceive and interpret musi…