1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
Video Question Answering Using CLIP-Guided Visual-Text Attention
Shuhong Ye, Weikai Kong, Chenglin Yao +2
Cross-modal learning of video and text plays a key role in Video Question Answering (VideoQA). In this paper, we propose a visual-text attention mechanism to utilize the Contrastiv…
cs.MM2023
Confidence-based Event-centric Online Video Question Answering on a Newly Constructed ATBS Dataset
Weikai Kong, Shuhong Ye, Chenglin Yao +1
Deep neural networks facilitate video question answering (VideoQA), but the real-world applications on video streams such as CCTV and live cast place higher demands on the solver.…