4 papers
Impact of Phonetics on Speaker Identity in Adversarial Voice Attack
Daniyal Kabir Dar, Qiben Yan, Li Xiao +1
Adversarial perturbations in speech pose a serious threat to automatic speech recognition (ASR) and speaker verification by introducing subtle waveform modifications that remain im…
Automated Assessment of Aesthetic Outcomes in Facial Plastic Surgery
Pegah Varghaei, Kiran Abraham-Aggarwal, Manoj T. Abraham +1
We introduce a scalable, interpretable computer-vision framework for quantifying aesthetic outcomes of facial plastic surgery using frontal photographs. Our pipeline leverages auto…
Foundation versus Domain-specific Models: Performance Comparison, Fusion, and Explainability in Face Recognition
Redwan Sony, Parisa Farmanifard, Arun Ross +1
In this paper, we address the following question: How do generic foundation models (e.g., CLIP, BLIP, GPT-4o, Grok-4) compare against a domain-specific face recognition model (viz.…
Benchmarking Foundation Models for Zero-Shot Biometric Tasks
Redwan Sony, Parisa Farmanifard, Hamzeh Alzwairy +2
The advent of foundation models, particularly Vision-Language Models (VLMs) and Multi-modal Large Language Models (MLLMs), has redefined the frontiers of artificial intelligence, e…