3 papers
cs.SD2026
Harmonica: Accurate and Lightweight Instrument-Agnostic Music Transcription
Longshen Ou, Héctor Martel, Joe Hennessy-Priest +1
This paper introduces Harmonica, a family of instrument-agnostic music transcription models built around multi-depth harmonic convolution. At each model scale, Harmonica achieves t…
eess.AS2026
Probing Low-Level Acoustic Attribute Encoding in CLAP Audio Embeddings
Héctor Martel, Joe Hennessy-Priest, Taemin Cho
Audio foundation models are widely adopted as general-purpose feature extractors, yet the internal structure of their learned representations remains insufficiently understood. In…
cs.SD2024
MR-MT3: Memory Retaining Multi-Track Music Transcription to Mitigate Instrument Leakage
Hao Hao Tan, Kin Wai Cheuk, Taemin Cho +2
This paper presents enhancements to the MT3 model, a state-of-the-art (SOTA) token-based multi-instrument automatic music transcription (AMT) model. Despite SOTA performance, MT3 h…