activity
20242026
most citedReading Between the Frames: Multi-Modal Depression Detection in Videos from Non-Verbal Cues

3 citations · 5 across the 7 of their papers we have counts for

collaborators

8 papers

eess.AS2026

CARE: A Multimodal Corpus for Studying Speech and Non-Verbal Communication Across Multiple Medical Conditions

David Gimeno-Gómez, Catarina Botelho, Carlos-D. Martínez-Hinarejos +2

Automatic analysis of multimodal speech has shown strong potential for computationally detecting and monitoring a wide range of neurological, psychiatric, and respiratory condition…

eess.AS2026

Cross-Modal Masking for Robust Silent Speech Synthesis Using sEMG and Lipreading

Eder del Blanco, David Gimeno-Gómez, Eva Navas +2

Speech restoration through silent speech interfaces (SSIs) has emerged as a promising assistive technology for individuals with impaired or absent laryngeal voice production. Among…

eess.AS2025

On the Relevance of Clinical Assessment Tasks for the Automatic Detection of Parkinson's Disease Medication State from Speech

David Gimeno-Gómez, Rubén Solera-Ureña, Anna Pompili +5

The automatic identification of medication states of Parkinson's disease (PD) patients can assist clinicians in monitoring and scheduling personalized treatments, as well as studyi…

eess.AS2024★ 1 cited

Tackling Cognitive Impairment Detection from Speech: A submission to the PROCESS Challenge

Catarina Botelho, David Gimeno-Gómez, Francisco Teixeira +11

This work describes our group's submission to the PROCESS Challenge 2024, with the goal of assessing cognitive decline through spontaneous speech, using three guided clinical tasks…

cs.CV2024

Tailored Design of Audio-Visual Speech Recognition Models using Branchformers

David Gimeno-Gómez, Carlos-D. Martínez-Hinarejos

Recent advances in Audio-Visual Speech Recognition (AVSR) have led to unprecedented achievements in the field, improving the robustness of this type of system in adverse, noisy env…

cs.CV2024★ 1 cited

AnnoTheia: A Semi-Automatic Annotation Toolkit for Audio-Visual Speech Technologies

José-M. Acosta-Triana, David Gimeno-Gómez, Carlos-D. Martínez-Hinarejos

More than 7,000 known languages are spoken around the world. However, due to the lack of annotated resources, only a small fraction of them are currently covered by speech technolo…