1 paper
Ariel Ephrat, Inbar Mosseri, Oran Lang +5
We present a joint audio-visual model for isolating a single speech signal from a mixture of sounds such as other speakers and background noise. Solving this task using only audio…