2 papers
cs.CV2018
Investigations on End-to-End Audiovisual Fusion
Michael Wand, Ngoc Thang Vu, Juergen Schmidhuber
Audiovisual speech recognition (AVSR) is a method to alleviate the adverse effect of noise in the acoustic signal. Leveraging recent developments in deep neural network-based speec…
cs.CV2017
Improving Speaker-Independent Lipreading with Domain-Adversarial Training
Michael Wand, Juergen Schmidhuber
We present a Lipreading system, i.e. a speech recognition system using only visual features, which uses domain-adversarial training for speaker independence. Domain-adversarial tra…