477 citations · 682 across the 55 of their papers we have counts for
8 papers · 1 filter
Multi-View Frame Reconstruction with Conditional GAN
Tahmida Mahmud, Mohammad Billah, Amit K. Roy-Chowdhury
Multi-view frame reconstruction is an important problem particularly when multiple frames are missing and past and future frames within the camera are far apart from the missing on…
Webly Supervised Joint Embedding for Cross-Modal Image-Text Retrieval
Niluthpol Chowdhury Mithun, Rameswar Panda, Evangelos E. Papalexakis +1
Cross-modal retrieval between visual data and natural language description remains a long-standing challenge in multimedia. While recent image-text retrieval methods offer great pr…
Incorporating Scalability in Unsupervised Spatio-Temporal Feature Learning
Sujoy Paul, Sourya Roy, Amit K. Roy-Chowdhury
Deep neural networks are efficient learning machines which leverage upon a large amount of manually labeled data for learning discriminative features. However, acquiring substantia…
Contemplating Visual Emotions: Understanding and Overcoming Dataset Bias
Rameswar Panda, Jianming Zhang, Haoxiang Li +3
While machine learning approaches to visual emotion recognition offer great promise, current methods consider training and testing models on small scale datasets covering limited v…
Adversarial Perturbations Against Real-Time Video Classification Systems
Shasha Li, Ajaya Neupane, Sujoy Paul +4
Recent research has demonstrated the brittleness of machine learning systems to adversarial perturbations. However, the studies have been mostly limited to perturbations on images…
W-TALC: Weakly-supervised Temporal Activity Localization and Classification
Sujoy Paul, Sourya Roy, Amit K Roy-Chowdhury
Most activity localization methods in the literature suffer from the burden of frame-wise annotation requirement. Learning from weak labels may be a potential solution towards redu…