3 papers
eess.AS2025
Adaptive Deterministic Flow Matching for Target Speaker Extraction
Tsun-An Hsieh, Minje Kim
Generative target speaker extraction (TSE) methods often produce more natural outputs than predictive models. Recent work based on diffusion or flow matching (FM) typically relies…
eess.AS2025
TGIF: Talker Group-Informed Familiarization of Target Speaker Extraction
Tsun-An Hsieh, Minje Kim
State-of-the-art target speaker extraction (TSE) systems are typically designed to generalize to any given mixing environment, necessitating a model with a large enough capacity as…
eess.AS2024
Multimodal Representation Loss Between Timed Text and Audio for Regularized Speech Separation
Tsun-An Hsieh, Heeyoul Choi, Minje Kim
Recent studies highlight the potential of textual modalities in conditioning the speech separation model's inference process. However, regularization-based methods remain underexpl…