3 papers
eess.AS2021
AMSS-Net: Audio Manipulation on User-Specified Sources with Textual Queries
Woosung Choi, Minseok Kim, Marco A. Martínez Ramírez +2
This paper proposes a neural network that performs audio transformations to user-specified sources (e.g., vocals) of a given audio track according to a given description while pres…
cs.SD2020
LaSAFT: Latent Source Attentive Frequency Transformation for Conditioned Source Separation
Woosung Choi, Minseok Kim, Jaehwa Chung +1
Recent deep-learning approaches have shown that Frequency Transformation (FT) blocks can significantly improve spectrogram-based single-source separation models by capturing freque…
eess.AS2019
Investigating U-Nets with various Intermediate Blocks for Spectrogram-based Singing Voice Separation
Woosung Choi, Minseok Kim, Jaehwa Chung +2
Singing Voice Separation (SVS) tries to separate singing voice from a given mixed musical signal. Recently, many U-Net-based models have been proposed for the SVS task, but there w…