3 papers
cs.SD2025
An Investigation of Incorporating Mamba for Speech Enhancement
Rong Chao, Wen-Huang Cheng, Moreno La Quatra +4
This work aims to investigate the use of a recently proposed, attention-free, scalable state-space model (SSM), Mamba, for the speech enhancement (SE) task. In particular, we emplo…
eess.AS2025
An Investigation on Combining Geometry and Consistency Constraints into Phase Estimation for Speech Enhancement
Chun-Wei Ho, Pin-Jui Ku, Hao Yen +3
We propose a novel iterative phase estimation framework, termed multi-source Griffin-Lim algorithm (MSGLA), for speech enhancement (SE) under additive noise conditions. The core id…
cs.CL2025
FlanEC: Exploring Flan-T5 for Post-ASR Error Correction
Moreno La Quatra, Valerio Mario Salerno, Yu Tsao +1
In this paper, we present an encoder-decoder model leveraging Flan-T5 for post-Automatic Speech Recognition (ASR) Generative Speech Error Correction (GenSEC), and we refer to it as…