3 papers
cs.SD2026
GDiffuSE: Diffusion-based speech enhancement with noise model guidance
Efrayim Yanir, David Burshtein, Sharon Gannot
This paper introduces a novel speech enhancement (SE) approach based on a denoising diffusion probabilistic model (DDPM), termed Guided diffusion for speech enhancement (GDiffuSE).…
eess.SP2025
(SP)-Net: A Neural Spatial Spectrum Method for DOA Estimation
Lioz Berman, Sharon Gannot, Tom Tirer
We consider the problem of estimating the directions of arrival (DOAs) of multiple sources from a single snapshot of an antenna array, a task with many practical applications. In s…
cs.CV2025
Video Editing for Audio-Visual Dubbing
Binyamin Manela, Sharon Gannot, Ethan Fetyaya
Visual dubbing, the synchronization of facial movements with new speech, is crucial for making content accessible across different languages, enabling broader global reach. However…