7 papers
On the de-duplication of the Lakh MIDI dataset
Eunjin Choi, Hyerin Kim, Jiwoo Ryu +2
A large-scale dataset is essential for training a well-generalized deep-learning model. Most such datasets are collected via scraping from various internet sources, inevitably intr…
Motive-level Analysis of Form-functions Association in Korean Folk song
Danbinaerin Han, Dasaem Jeong, Juhan Nam
Computational analysis of folk song audio is challenging due to structural irregularities and the need for manual annotation. We propose a method for automatic motive segmentation…
Can Audio Reveal Music Performance Difficulty? Insights from the Piano Syllabus Dataset
Pedro Ramoneda, Minhee Lee, Dasaem Jeong +2
Automatically estimating the performance difficulty of a music piece represents a key process in music education to create tailored curricula according to the individual needs of t…
Enriching Music Descriptions with a Finetuned-LLM and Metadata for Text-to-Music Retrieval
SeungHeon Doh, Minhee Lee, Dasaem Jeong +1
Text-to-Music Retrieval, finding music based on a given natural language query, plays a pivotal role in content discovery within extensive music databases. To address this challeng…
K-pop Lyric Translation: Dataset, Analysis, and Neural-Modelling
Haven Kim, Jongmin Jung, Dasaem Jeong +1
Lyric translation, a field studied for over a century, is now attracting computational linguistics researchers. We identified two limitations in previous studies. Firstly, lyric tr…
Musical Word Embedding for Music Tagging and Retrieval
SeungHeon Doh, Jongpil Lee, Dasaem Jeong +1
Word embedding has become an essential means for text-based information retrieval. Typically, word embeddings are learned from large quantities of general and unstructured text dat…