activity
20192022
collaborators

5 papers

eess.AS2022

Device-Directed Speech Detection: Regularization via Distillation for Weakly-Supervised Models

Vineet Garg, Ognjen Rudovic, Pranay Dighe +5

We address the problem of detecting speech directed to a device that does not contain a specific wake-word. Specifically, we focus on audio coming from a touch-based invocation. Mi…

eess.AS2020

Modality Dropout for Improved Performance-driven Talking Faces

Ahmed Hussen Abdelaziz, Barry-John Theobald, Paul Dixon +3

We describe our novel deep learning approach for driving animated faces using both acoustic and visual information. In particular, speech-related facial movements are generated usi…

cs.LG2020

On the Role of Visual Cues in Audiovisual Speech Enhancement

Zakaria Aldeneh, Anushree Prasanna Kumar, Barry-John Theobald +4

We present an introspection of an audiovisual speech enhancement model. In particular, we focus on interpreting how a neural audiovisual speech enhancement model uses visual cues t…

eess.AS2019

On Neural Phone Recognition of Mixed-Source ECoG Signals

Ahmed Hussen Abdelaziz, Shuo-Yiin Chang, Nelson Morgan +5

The emerging field of neural speech recognition (NSR) using electrocorticography has recently attracted remarkable research interest for studying how human brains recognize speech…

eess.AS2019

Speaker-Independent Speech-Driven Visual Speech Synthesis using Domain-Adapted Acoustic Models

Ahmed Hussen Abdelaziz, Barry-John Theobald, Justin Binder +5

Speech-driven visual speech synthesis involves mapping features extracted from acoustic speech to the corresponding lip animation controls for a face model. This mapping can take m…