2 citations · 3 across the 4 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2026
Decoupled Latent Flow Matching for Few-Step Joint Vocal-Accompaniment Separation
Lishi Zuo, Youzhi Tu, Lu Yi +4
Generative modeling provides a flexible way to model mixture-conditioned source distributions, but iterative diffusion and flow matching models are costly for long music signals. T…
cs.SD2023★ 1 cited
Phonetic-aware speaker embedding for far-field speaker verification
Zezhong Jin, Youzhi Tu, Man-Wai Mak
When a speaker verification (SV) system operates far from the sound sourced, significant challenges arise due to the interference of noise and reverberation. Studies have shown tha…
cs.SD2023★ 2 cited
Self-supervised Neural Factor Analysis for Disentangling Utterance-level Speech Representations
Weiwei Lin, Chenhang He, Man-Wai Mak +1
Self-supervised learning (SSL) speech models such as wav2vec and HuBERT have demonstrated state-of-the-art performance on automatic speech recognition (ASR) and proved to be extrem…