185 citations · 544 across the 17 of their papers we have counts for
27 papers
Comparative Study of Domain-adapted VLMs for General Document Visual Question Answering
Miguel Lopez-Duran, Elena Marrero, Julian Fierrez +8
Document Visual Question Answering (DocVQA) presents a complex multimodal challenge, requiring models to exploit visual, textual, and layout information from documents. Although Vi…
Exploring Vision-Language Models for Online Signature Verification: A Zero-Shot Capability Study
Marta Robledo-Moreno, Ruben Vera-Rodriguez, Ruben Tolosana +1
Recent advancements in Vision-Language Models (VLMs) have demonstrated strong capabilities in general visual reasoning, yet their applicability to rigorous biometric tasks remains…
Active Membership Inference Test (aMINT): Enhancing Model Auditability with Multi-Task Learning
Daniel DeAlcala, Aythami Morales, Julian Fierrez +3
Active Membership Inference Test (aMINT) is a method designed to detect whether given data were used during the training of machine learning models. In Active MINT, we propose a no…
AirSignatureDB: Exploring In-Air Signature Biometrics in the Wild and its Privacy Concerns
Marta Robledo-Moreno, Ruben Vera-Rodriguez, Ruben Tolosana +3
Behavioral biometrics based on smartphone motion sensors are growing in popularity for authentication purposes. In this study, AirSignatureDB is presented: a new publicly accessibl…
AG-VPReID 2025: Aerial-Ground Video-based Person Re-identification Challenge Results
Kien Nguyen, Clinton Fookes, Sridha Sridharan +20
Person re-identification (ReID) across aerial and ground vantage points has become crucial for large-scale surveillance and public safety applications. Although significant progres…
Are Vision-Language Models Ready for Dietary Assessment? Exploring the Next Frontier in AI-Powered Food Image Recognition
Sergio Romero-Tapiador, Ruben Tolosana, Blanca Lacruz-Pleguezuelos +7
Automatic dietary assessment based on food images remains a challenge, requiring precise food detection, segmentation, and classification. Vision-Language Models (VLMs) offer new p…