4 papers
MLLM-based Textual Explanations for Face Comparison
Redwan Sony, Anil K Jain, Arun Ross
Multimodal Large Language Models (MLLMs) have recently been proposed as a means to generate natural-language explanations for face recognition decisions. While such explanations fa…
Foundation versus Domain-specific Models: Performance Comparison, Fusion, and Explainability in Face Recognition
Redwan Sony, Parisa Farmanifard, Arun Ross +1
In this paper, we address the following question: How do generic foundation models (e.g., CLIP, BLIP, GPT-4o, Grok-4) compare against a domain-specific face recognition model (viz.…
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
Redwan Sony, Parisa Farmanifard, Hamzeh Alzwairy +2
The advent of foundation models, particularly Vision-Language Models (VLMs) and Multi-modal Large Language Models (MLLMs), has redefined the frontiers of artificial intelligence, e…
A Parametric Approach to Adversarial Augmentation for Cross-Domain Iris Presentation Attack Detection
Debasmita Pal, Redwan Sony, Arun Ross
Iris-based biometric systems are vulnerable to presentation attacks (PAs), where adversaries present physical artifacts (e.g., printed iris images, textured contact lenses) to defe…