2 papers
cs.CV2026
Quran-MD: A Fine-Grained Multilingual Multimodal Dataset of the Quran
Muhammad Umar Salman, Mohammad Areeb Qazi, Mohammed Talha Alam
We present Quran MD, a comprehensive multimodal dataset of the Quran that integrates textual, linguistic, and audio dimensions at the verse and word levels. For each verse (ayah),…
cs.CV2025
Robust and Calibrated Detection of Authentic Multimedia Content
Sarim Hashmi, Abdelrahman Elsayed, Mohammed Talha Alam +2
Generative models can synthesize highly realistic content, so-called deepfakes, that are already being misused at scale to undermine digital media authenticity. Current deepfake de…