16 citations · 41 across the 6 of their papers we have counts for
5 papers · 1 filter
X-Pool: Cross-Modal Language-Video Attention for Text-Video Retrieval
Satya Krishna Gorti, Noel Vouitsis, Junwei Ma +4
In text-video retrieval, the objective is to learn a cross-modal similarity function between a text and a video that ranks relevant text-video pairs higher than irrelevant pairs. H…
Weakly Supervised Action Selection Learning in Video
Junwei Ma, Satya Krishna Gorti, Maksims Volkovs +1
Localizing actions in video is a core task in computer vision. The weakly supervised temporal localization problem investigates whether this task can be adequately solved with only…
Learning Effective Visual Relationship Detector on 1 GPU
Yichao Lu, Cheng Chang, Himanshu Rai +2
We present our winning solution to the Open Images 2019 Visual Relationship challenge. This is the largest challenge of its kind to date with nearly 9 million training images. Chal…
Cross-Class Relevance Learning for Temporal Concept Localization
Junwei Ma, Satya Krishna Gorti, Maksims Volkovs +2
We present a novel Cross-Class Relevance Learning approach for the task of temporal concept localization. Most localization architectures rely on feature extraction layers followed…
Semi-Supervised Exploration in Image Retrieval
Cheng Chang, Himanshu Rai, Satya Krishna Gorti +4
We present our solution to Landmark Image Retrieval Challenge 2019. This challenge was based on the large Google Landmarks Dataset V2[9]. The goal was to retrieve all database imag…