5 papers
Face-Voice Association with Inductive Bias for Maximum Class Separation
Marta Moscati, Oleksandr Kats, Mubashir Noman +4
Face-voice association is widely studied in multimodal learning and is approached representing faces and voices with embeddings that are close for a same person and well separated…
Linking Faces and Voices Across Languages: Insights from the FAME 2026 Challenge
Marta Moscati, Ahmed Abdullah, Muhammad Saad Saeed +7
Over half of the world's population is bilingual and people often communicate under multilingual scenarios. The Face-Voice Association in Multilingual Environments (FAME) 2026 Chal…
RobustA: Robust Anomaly Detection in Multimodal Data
Salem AlMarri, Muhammad Irzam Liaqat, Muhammad Zaigham Zaheer +3
In recent years, multimodal anomaly detection methods have demonstrated remarkable performance improvements over video-only models. However, real-world multimodal data is often cor…
Face-voice Association in Multilingual Environments (FAME) 2026 Challenge Evaluation Plan
Marta Moscati, Ahmed Abdullah, Muhammad Saad Saeed +7
The advancements of technology have led to the use of multimodal systems in various real-world applications. Among them, audio-visual systems are among the most widely used multimo…
GenMix: Effective Data Augmentation with Generative Diffusion Model Image Editing
Khawar Islam, Muhammad Zaigham Zaheer, Arif Mahmood +2
Data augmentation is widely used to enhance generalization in visual classification tasks. However, traditional methods struggle when source and target domains differ, as in domain…