collaborators

6 papers

cs.CV2025

Enhancing Action Recognition by Leveraging the Hierarchical Structure of Actions and Textual Context

Manuel Benavent-Lledo, David Mulero-Pérez, David Ortiz-Perez +2

We propose a novel approach to improve action recognition by exploiting the hierarchical organization of actions and by incorporating contextualized textual information, including…

cs.LG2025

CogniAlign: Word-Level Multimodal Speech Alignment with Gated Cross-Attention for Alzheimer's Detection

David Ortiz-Perez, Manuel Benavent-Lledo, Javier Rodriguez-Juan +2

Early detection of cognitive disorders such as Alzheimer's disease is critical for enabling timely clinical intervention and improving patient outcomes. In this work, we introduce…

cs.LG2025

Deep Insights into Cognitive Decline: A Survey of Leveraging Non-Intrusive Modalities with Deep Learning Techniques

David Ortiz-Perez, Manuel Benavent-Lledo, Jose Garcia-Rodriguez +2

Cognitive decline is a natural part of aging. However, under some circumstances, this decline is more pronounced than expected, typically due to disorders such as Alzheimer's disea…

cs.CV2025

Text-driven Online Action Detection

Manuel Benavent-Lledo, David Mulero-Pérez, David Ortiz-Perez +1

Detecting actions as they occur is essential for applications like video surveillance, autonomous driving, and human-robot interaction. Known as online action detection, this task…

cs.CV2025

Visual WetlandBirds Dataset: Bird Species Identification and Behavior Recognition in Videos

Javier Rodriguez-Juan, David Ortiz-Perez, Manuel Benavent-Lledo +5

The current biodiversity loss crisis makes animal monitoring a relevant field of study. In light of this, data collected through monitoring can provide essential insights, and info…

cs.CV2024

Detecting Facial Image Manipulations with Multi-Layer CNN Models

Alejandro Marco Montejano, Angela Sanchez Perez, Javier Barrachina +3

The rapid evolution of digital image manipulation techniques poses significant challenges for content verification, with models such as stable diffusion and mid-journey producing h…