8 citations · 20 across the 5 of their papers we have counts for
4 papers · 1 filter
Visual Speech Enhancement
Aviv Gabbay, Asaph Shamir, Shmuel Peleg
When video is shot in noisy environment, the voice of a speaker seen in the video can be enhanced using the visible mouth movements, reducing background noise. While most existing…
Improved Speech Reconstruction from Silent Video
Ariel Ephrat, Tavi Halperin, Shmuel Peleg
Speechreading is the task of inferring phonetic information from visually observed articulatory facial movements, and is a notoriously difficult task for humans to perform. In this…
Seeing Through Noise: Visually Driven Speaker Separation and Enhancement
Aviv Gabbay, Ariel Ephrat, Tavi Halperin +1
Isolating the voice of a specific person while filtering out other voices or background noises is challenging when video is shot in noisy environments. We propose audio-visual meth…
Vid2speech: Speech Reconstruction from Silent Video
Ariel Ephrat, Shmuel Peleg
Speechreading is a notoriously difficult task for humans to perform. In this paper we present an end-to-end model based on a convolutional neural network (CNN) for generating an in…