3 papers
cs.SD2026
TF-MossFormer: Integrating Convolution Gated Local-Global Attentions for Enhanced Time-Frequency Domain Monaural Speech Separation
Shengkui Zhao, Zexu Pan, Haoxu Wang +3
Transformers with global attention capture long-range dependencies but can miss the fine-grained local continuity crucial for speech separation. We propose TF-MossFormer, a time-fr…
cs.SD2026
E2E-AEC: Implementing an end-to-end neural network learning approach for acoustic echo cancellation
Yiheng Jiang, Biao Tian, Haoxu Wang +4
We propose a novel neural network-based end-to-end acoustic echo cancellation (E2E-AEC) method capable of streaming inference, which operates effectively without reliance on tradit…
eess.AS2026
FlowSE-GRPO: Training Flow Matching Speech Enhancement via Online Reinforcement Learning
Haoxu Wang, Biao Tian, Yiheng Jiang +5
Generative speech enhancement offers a promising alternative to traditional discriminative methods by modeling the distribution of clean speech conditioned on noisy inputs. Post-tr…