collaborators

6 papers

cs.SD2025

Unified Cross-modal Translation of Score Images, Symbolic Music, and Performance Audio

Jongmin Jung, Dongmin Kim, Sihun Lee +5

Music exists in various modalities, such as score images, symbolic scores, MIDI, and audio. Translations between each modality are established as core tasks of music information re…

cs.SD2025

LAV: Audio-Driven Dynamic Visual Generation with Neural Compression and StyleGAN2

Jongmin Jung, Dasaem Jeong

This paper introduces LAV (Latent Audio-Visual), a system that integrates EnCodec's neural audio compression with StyleGAN2's generative capabilities to produce visually dynamic ou…

cs.SD2025

Boundary Regression for Leitmotif Detection in Music Audio

Sihun Lee, Dasaem Jeong

Leitmotifs are musical phrases that are reprised in various forms throughout a piece. Due to diverse variations and instrumentation, detecting the occurrence of leitmotifs from aud…

cs.SD2024

MusicGen-Chord: Advancing Music Generation through Chord Progressions and Interactive Web-UI

Jongmin Jung, Andreas Jansson, Dasaem Jeong

MusicGen is a music generation language model (LM) that can be conditioned on textual descriptions and melodic features. We introduce MusicGen-Chord, which extends this capability…

cs.SD2024

Towards Computational Analysis of Pansori Singing

Sangheon Park, Danbinaerin Han, Dasaem Jeong

Pansori is one of the most representative vocal genres of Korean traditional music, which has an elaborated vocal melody line with strong vibrato. Although the music is transmitted…

cs.SD2024

ViolinDiff: Enhancing Expressive Violin Synthesis with Pitch Bend Conditioning

Daewoong Kim, Hao-Wen Dong, Dasaem Jeong

Modeling the natural contour of fundamental frequency (F0) plays a critical role in music audio synthesis. However, transcribing and managing multiple F0 contours in polyphonic mus…