Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
Assessing Factual Music Comprehension in Large Audio Language Models
Daniel Chenyu Lin, Michael Freeman, John Thickstun
Large audio language models (LALMs) leverage multimodal representations to generate open-ended answers to natural language queries about audio. In this paper, we (1) provide empiri…
cs.SD2024
Do Music Generation Models Encode Music Theory?
Megan Wei, Michael Freeman, Chris Donahue +1
Music foundation models possess impressive music generation capabilities. When people compose music, they may infuse their understanding of music into their work, by using notes an…