4 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 4 cited
Modeling Semantic Composition with Syntactic Hypergraph for Video Question Answering
Zenan Xu, Wanjun Zhong, Qinliang Su +2
A key challenge in video question answering is how to realize the cross-modal semantic alignment between textual concepts and corresponding visual objects. Existing methods mostly…
cs.CV2021★ 2 cited
The Multi-Modal Video Reasoning and Analyzing Competition
Haoran Peng, He Huang, Li Xu +15
In this paper, we introduce the Multi-Modal Video Reasoning and Analyzing Competition (MMVRAC) workshop in conjunction with ICCV 2021. This competition is composed of four differen…