7 papers · 1 filter
Aligning Generative Music AI with Human Preferences: Methods and Challenges
Dorien Herremans, Abhinaba Roy
Recent advances in generative AI for music have achieved remarkable fidelity and stylistic diversity, yet these systems often fail to align with nuanced human preferences due to th…
SonicVerse: Multi-Task Learning for Music Feature-Informed Captioning
Anuradha Chopra, Abhinaba Roy, Dorien Herremans
Detailed captions that accurately reflect the characteristics of a music piece can enrich music databases and drive forward research in music AI. This paper introduces a multi-task…
Text2midi-InferAlign: Improving Symbolic Music Generation with Inference-Time Alignment
Abhinaba Roy, Geeta Puri, Dorien Herremans
We present Text2midi-InferAlign, a novel technique for improving symbolic music generation at inference time. Our method leverages text-to-audio alignment and music structural alig…
MelodySim: Measuring Melody-aware Music Similarity for Plagiarism Detection
Tongyu Lu, Charlotta-Marlena Geist, Jan Melechovsky +2
We propose MelodySim, a melody-aware music similarity model and dataset for plagiarism detection. First, we introduce a novel method to construct a dataset focused on melodic simil…
JamendoMaxCaps: A Large Scale Music-caption Dataset with Imputed Metadata
Abhinaba Roy, Renhang Liu, Tongyu Lu +1
We introduce JamendoMaxCaps, a large-scale music-caption dataset featuring over 362,000 freely licensed instrumental tracks from the renowned Jamendo platform. The dataset includes…
Text2midi: Generating Symbolic Music from Captions
Keshav Bhandari, Abhinaba Roy, Kyra Wang +3
This paper introduces text2midi, an end-to-end model to generate MIDI files from textual descriptions. Leveraging the growing popularity of multimodal generative approaches, text2m…