41 citations · 44 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 3 cited
Locate before Answering: Answer Guided Question Localization for Video Question Answering
Tianwen Qian, Ran Cui, Jingjing Chen +3
Video question answering (VideoQA) is an essential task in vision-language understanding, which has attracted numerous research attention recently. Nevertheless, existing works mos…
cs.CV2022★ 41 cited
Video Moment Retrieval from Text Queries via Single Frame Annotation
Ran Cui, Tianwen Qian, Pai Peng +5
Video moment retrieval aims at finding the start and end timestamps of a moment (part of a video) described by a given natural language query. Fully supervised methods need complet…