From the 1 of 6 linked papers with an AI index.
6 papers
Learning Speaker Identity Beyond Language and Modality Constraints: Insights from the POLY-SIM 2026 Challenge
Marta Moscati, Muhammad Saad Saeed, Marina Zanoni +9
The paper describes the POLY-SIM 2026 challenge, which focuses on developing multimodal speaker identification systems that remain robust when audio or visual data are missing and…
Iterative Definition Refinement for Zero-Shot Classification via LLM-Based Semantic Prototype Optimization
Naeem Rehmat, Muhammad Saad Saeed, Ijaz Ul Haq +1
Web filtering systems rely on accurate web content classification to block cyber threats, prevent data exfiltration, and ensure compliance. However, classification is increasingly…
POLY-SIM: Polyglot Speaker Identification with Missing Modality Grand Challenge 2026 Evaluation Plan
Marta Moscati, Muhammad Saad Saeed, Marina Zanoni +8
Multimodal speaker identification systems typically assume the availability of complete and homogeneous audio-visual modalities during both training and testing. However, in real-w…
Linking Faces and Voices Across Languages: Insights from the FAME 2026 Challenge
Marta Moscati, Ahmed Abdullah, Muhammad Saad Saeed +7
Over half of the world's population is bilingual and people often communicate under multilingual scenarios. The Face-Voice Association in Multilingual Environments (FAME) 2026 Chal…
Face-voice Association in Multilingual Environments (FAME) 2026 Challenge Evaluation Plan
Marta Moscati, Ahmed Abdullah, Muhammad Saad Saeed +7
The advancements of technology have led to the use of multimodal systems in various real-world applications. Among them, audio-visual systems are among the most widely used multimo…
Realism to Deception: Investigating Deepfake Detectors Against Face Enhancement
Muhammad Saad Saeed, Ijaz Ul Haq, Khalid Malik
Face enhancement techniques are widely used to enhance facial appearance. However, they can inadvertently distort biometric features, leading to significant decrease in the accurac…