2 papers
cs.SD2025
AUREXA-SE: Audio-Visual Unified Representation Exchange Architecture with Cross-Attention and Squeezeformer for Speech Enhancement
M. Sajid, Deepanshu Gupta, Yash Modi +7
In this paper, we propose AUREXA-SE (Audio-Visual Unified Representation Exchange Architecture with Cross-Attention and Squeezeformer for Speech Enhancement), a progressive bimodal…
cs.SD2024
LSTMSE-Net: Long Short Term Speech Enhancement Network for Audio-visual Speech Enhancement
Arnav Jain, Jasmer Singh Sanjotra, Harshvardhan Choudhary +6
In this paper, we propose long short term memory speech enhancement network (LSTMSE-Net), an audio-visual speech enhancement (AVSE) method. This innovative method leverages the com…