3 papers
eess.AS2026
MeCo: One-Step MeanFlow-based Corrector for Multi-Channel Speech Separation
Dohwan Kim, Jung-Woo Choi
While discriminative models for multi-channel speech separation excel in reference-based metrics, they often exhibit suboptimal human listening quality. To address this, we propose…
eess.AS2025
SoundCompass: Navigating Target Sound Extraction With Effective Directional Clue Integration In Complex Acoustic Scenes
Dayun Choi, Jung-Woo Choi
Recent advances in target sound extraction (TSE) utilize directional clues derived from direction of arrival (DoA), which represent an inherent spatial property of sound available…
eess.AS2025
CST-former: Multidimensional Attention-based Transformer for Sound Event Localization and Detection in Real Scenes
Yusun Shul, Dayun Choi, Jung-Woo Choi
Sound event localization and detection (SELD) is a task for the classification of sound events and the identification of direction of arrival (DoA) utilizing multichannel acoustic…