activity
20182022
most citedHow to Teach DNNs to Pay Attention to the Visual Modality in Speech Recognition

39 citations · 40 across the 4 of their papers we have counts for

collaborators

11 papers

eess.AS20221 cited

Learnable Acoustic Frontends in Bird Activity Detection

Mark Anderson, Naomi Harte

Autonomous recording units and passive acoustic monitoring present minimally intrusive methods of collecting bioacoustics data. Combining this data with species agnostic bird activ…

eess.AS2020

AV Taris: Online Audio-Visual Speech Recognition

George Sterpu, Naomi Harte

In recent years, Automatic Speech Recognition (ASR) technology has approached human-level performance on conversational speech under relatively clean listening conditions. In more…

eess.AS2020

Learning to Count Words in Fluent Speech enables Online Speech Recognition

George Sterpu, Christian Saam, Naomi Harte

Sequence to Sequence models, in particular the Transformer, achieve state of the art results in Automatic Speech Recognition. Practical usage is however limited to cases where full…

eess.AS2020

Should we hard-code the recurrence concept or learn it instead ? Exploring the Transformer architecture for Audio-Visual Speech Recognition

George Sterpu, Christian Saam, Naomi Harte

The audio-visual speech fusion strategy AV Align has shown significant performance improvements in audio-visual speech recognition (AVSR) on the challenging LRS2 dataset. Performan…

cs.CL2020

Neural Generation of Dialogue Response Timings

Matthew Roddy, Naomi Harte

The timings of spoken response offsets in human dialogue have been shown to vary based on contextual elements of the dialogue. We propose neural models that simulate the distributi…

eess.AS202039 cited

How to Teach DNNs to Pay Attention to the Visual Modality in Speech Recognition

George Sterpu, Christian Saam, Naomi Harte

Audio-Visual Speech Recognition (AVSR) seeks to model, and thereby exploit, the dynamic relationship between a human voice and the corresponding mouth movements. A recently propose…