12 citations · 12 across the 1 of their papers we have counts for
6 papers
Character Matters: Video Story Understanding with Character-Aware Relations
Shijie Geng, Ji Zhang, Zuohui Fu +3
Different from short videos and GIFs, video stories contain clear plots and lists of principal characters. Without identifying the connection between appearing people and character…
2nd Place Solution to the GQA Challenge 2019
Shijie Geng, Ji Zhang, Hang Zhang +2
We present a simple method that achieves unexpectedly superior performance for Complex Reasoning involved Visual Question Answering. Our solution collects statistical features from…
Graphical Contrastive Losses for Scene Graph Parsing
Ji Zhang, Kevin J. Shih, Ahmed Elgammal +2
Most scene graph parsers use a two-stage pipeline to detect visual relationships: the first stage detects entities, and the second predicts the predicate for each entity pair using…
An Interpretable Model for Scene Graph Generation
Ji Zhang, Kevin Shih, Andrew Tao +2
We propose an efficient and interpretable scene graph generator. We consider three types of features: visual, spatial and semantic, and we use a late fusion strategy such that each…
Introduction to the 1st Place Winning Model of OpenImages Relationship Detection Challenge
Ji Zhang, Kevin Shih, Andrew Tao +2
This article describes the model we built that achieved 1st place in the OpenImage Visual Relationship Detection Challenge on Kaggle. Three key factors contribute the most to our s…
Large-Scale Visual Relationship Understanding
Ji Zhang, Yannis Kalantidis, Marcus Rohrbach +3
Large scale visual understanding is challenging, as it requires a model to handle the widely-spread and imbalanced distribution of <subject, relation, object> triples. In real-worl…