3 citations · 4 across the 3 of their papers we have counts for
1 paper · 1 filter
Julius Richter, Simone Frintrop, Timo Gerkmann
This paper introduces an audio-visual speech enhancement system that leverages score-based generative models, also known as diffusion models, conditioned on visual information. In…