51 citations · 151 across the 25 of their papers we have counts for
3 papers · 2 filters
Constructing a Visual Relationship Authenticity Dataset
Chenhui Chu, Yuto Takebayashi, Mishra Vipul +1
A visual relationship denotes a relationship between two objects in an image, which can be represented as a triplet of (subject; predicate; object). Visual relationship detection i…
A Dataset and Baselines for Visual Question Answering on Art
Noa Garcia, Chentao Ye, Zihua Liu +5
Answering questions related to art pieces (paintings) is a difficult task, as it implies the understanding of not only the visual information that is shown in the picture, but also…
Knowledge-Based Visual Question Answering in Videos
Noa Garcia, Mayu Otani, Chenhui Chu +1
We propose a novel video understanding task by fusing knowledge-based and video question answering. First, we introduce KnowIT VQA, a video dataset with 24,282 human-generated ques…