diffusion models 1hierarchical guidance 1in-context learning 1timbre synthesis 1zero-shot instrument cloning 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.SD2026
P-MUSE: Prompt-MIDI-Optional Model for Unified Instrumental Music Synthesis and Editing
Chong Jing, Junan Zhang, Jing Yang +3
MIDI-to-Music system renders the melody and rhythm of a target MIDI sequence into musical segment while cloning instrument timbre from a prompt recording. Existing systems typicall…
cs.SD2026
Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance
Chong Jing, Junan Zhang, Jing Yang +3
The paper presents Anysynth, a diffusion‑transformer synthesizer that can render arbitrary target MIDI sequences with the timbre of an unseen instrument by directly conditioning on…
cs.SD2026
EigeNet: Geometry-Informed Multi-Modal Learning for Few-shot Novel View RIR Prediction
Chong Jing, Zitong Lan, Junan Zhang +1
Predicting spatially varying Room Impulse Response (RIR) from sparse observations is a critical but highly challenging inverse problem for immersive spatial audio rendering. In thi…