11 papers
Deep Latent Space Learning for Cross-modal Mapping of Audio and Visual Signals
Shah Nawaz, Muhammad Kamran Janjua, Ignazio Gallo +2
We propose a novel deep training algorithm for joint representation of audio and visual information which consists of a single stream network (SSNet) coupled with a novel loss func…
Picture What you Read
Ignazio Gallo, Shah Nawaz, Alessandro Calefati +2
Visualization refers to our ability to create an image in our head based on the text we read or the words we hear. It is one of the many skills that makes reading comprehension pos…
A Classification Methodology based on Subspace Graphs Learning
Riccardo La Grassa, Ignazio Gallo, Alessandro Calefati +1
In this paper, we propose a design methodology for one-class classifiers using an ensemble-of-classifiers approach. The objective is to select the best structures created during th…
Do Cross Modal Systems Leverage Semantic Relationships?
Shah Nawaz, Muhammad Kamran Janjua, Ignazio Gallo +3
Current cross-modal retrieval systems are evaluated using R@K measure which does not leverage semantic relationships rather strictly follows the manually marked image text query pa…
Binary Classification using Pairs of Minimum Spanning Trees or N-ary Trees
Riccardo La Grassa, Ignazio Gallo, Alessandro Calefati +1
One-class classifiers are trained with target class only samples. Intuitively, their conservative modelling of the class description may benefit classical classification tasks wher…
Aiding Intra-Text Representations with Visual Context for Multimodal Named Entity Recognition
Omer Arshad, Ignazio Gallo, Shah Nawaz +1
With massive explosion of social media such as Twitter and Instagram, people daily share billions of multimedia posts, containing images and text. Typically, text in these posts is…