4 papers
VisG AV-HuBERT: Viseme-Guided AV-HuBERT
Aristeidis Papadopoulos, Rishabh Jain, Naomi Harte
Audio-Visual Speech Recognition (AVSR) systems nowadays integrate Large Language Model (LLM) decoders with transformer-based encoders, achieving state-of-the-art results. However,…
Interpreting the Role of Visemes in Audio-Visual Speech Recognition
Aristeidis Papadopoulos, Naomi Harte
Audio-Visual Speech Recognition (AVSR) models have surpassed their audio-only counterparts in terms of performance. However, the interpretability of AVSR systems, particularly the…
Uncovering the Visual Contribution in Audio-Visual Speech Recognition
Zhaofeng Lin, Naomi Harte
Audio-Visual Speech Recognition (AVSR) combines auditory and visual speech cues to enhance the accuracy and robustness of speech recognition systems. Recent advancements in AVSR ha…
Noise-Robust Hearing Aid Voice Control
Iván López-Espejo, Eros Roselló, Amin Edraki +2
Advancing the design of robust hearing aid (HA) voice control is crucial to increase the HA use rate among hard of hearing people as well as to improve HA users' experience. In thi…