10 citations · 10 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 10 cited
NExT-QA:Next Phase of Question-Answering to Explaining Temporal Actions
Junbin Xiao, Xindi Shang, Angela Yao +1
We introduce NExT-QA, a rigorously designed video question answering (VideoQA) benchmark to advance video understanding from describing to explaining the temporal actions. Based on…
cs.CV2020
Visual Relation Grounding in Videos
Junbin Xiao, Xindi Shang, Xun Yang +2
In this paper, we explore a novel task named visual Relation Grounding in Videos (vRGV). The task aims at spatio-temporally localizing the given relations in the form of subject-pr…