1 citations · 2 across the 4 of their papers we have counts for
4 papers
Single-branch Network for Multimodal Training
Muhammad Saad Saeed, Shah Nawaz, Muhammad Haris Khan +4
With the rapid growth of social media platforms, users are sharing billions of multimedia posts containing audio, images, and text. Researchers have focused on building autonomous…
Speaker Recognition in Realistic Scenario Using Multimodal Data
Saqlain Hussain Shah, Muhammad Saad Saeed, Shah Nawaz +1
In recent years, an association is established between faces and voices of celebrities leveraging large scale audio-visual information from YouTube. The availability of large scale…
Learning Branched Fusion and Orthogonal Projection for Face-Voice Association
Muhammad Saad Saeed, Shah Nawaz, Muhammad Haris Khan +3
Recent years have seen an increased interest in establishing association between faces and voices of celebrities leveraging audio-visual information from YouTube. Prior works adopt…
Fusion and Orthogonal Projection for Improved Face-Voice Association
Muhammad Saad Saeed, Muhammad Haris Khan, Shah Nawaz +2
We study the problem of learning association between face and voice, which is gaining interest in the computer vision community lately. Prior works adopt pairwise or triplet loss f…