3 papers
eess.AS2025
SAGA-SR: Semantically and Acoustically Guided Audio Super-Resolution
Jaekwon Im, Juhan Nam
Versatile audio super-resolution (SR) aims to predict high-frequency components from low-resolution audio across diverse domains such as speech, music, and sound effects. Existing…
eess.AS2025
FlashSR: One-step Versatile Audio Super-resolution via Diffusion Distillation
Jaekwon Im, Juhan Nam
Versatile audio super-resolution (SR) is the challenging task of restoring high-frequency components from low-resolution audio with sampling rates between 4kHz and 32kHz in various…
cs.SD2022
Neural Vocoder Feature Estimation for Dry Singing Voice Separation
Jaekwon Im, Soonbeom Choi, Sangeon Yong +1
Singing voice separation (SVS) is a task that separates singing voice audio from its mixture with instrumental audio. Previous SVS studies have mainly employed the spectrogram mask…