115 citations · 127 across the 3 of their papers we have counts for
11 papers
Masked Linear Regression for Learning Local Receptive Fields for Facial Expression Synthesis
Nazar Khan, Arbish Akram, Arif Mahmood +2
Compared to facial expression recognition, expression synthesis requires a very high-dimensional mapping. This problem exacerbates with increasing image sizes and limits existing e…
A Self-Reasoning Framework for Anomaly Detection Using Video-Level Labels
Muhammad Zaigham Zaheer, Arif Mahmood, Hochul Shin +1
Anomalous event detection in surveillance videos is a challenging and practical research problem among image and video processing community. Compared to the frame-level annotations…
Cross-modal Speaker Verification and Recognition: A Multilingual Perspective
Muhammad Saad Saeed, Shah Nawaz, Pietro Morerio +4
Recent years have seen a surge in finding association between faces and voices within a cross-modal biometric application along with speaker recognition. Inspired from this, we int…
Deep Latent Space Learning for Cross-modal Mapping of Audio and Visual Signals
Shah Nawaz, Muhammad Kamran Janjua, Ignazio Gallo +2
We propose a novel deep training algorithm for joint representation of audio and visual information which consists of a single stream network (SSNet) coupled with a novel loss func…
Do Cross Modal Systems Leverage Semantic Relationships?
Shah Nawaz, Muhammad Kamran Janjua, Ignazio Gallo +3
Current cross-modal retrieval systems are evaluated using R@K measure which does not leverage semantic relationships rather strictly follows the manually marked image text query pa…
Leveraging Orientation for Weakly Supervised Object Detection with Application to Firearm Localization
Javed Iqbal, Muhammad Akhtar Munir, Arif Mahmood +2
Automatic detection of firearms is important for enhancing the security and safety of people, however, it is a challenging task owing to the wide variations in shape, size, and app…