4 citations · 5 across the 11 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
Multi-Clue Reasoning with Memory Augmentation for Knowledge-based Visual Question Answering
Chengxiang Yin, Zhengping Che, Kun Wu +2
Visual Question Answering (VQA) has emerged as one of the most challenging tasks in artificial intelligence due to its multi-modal nature. However, most existing VQA methods are in…
cs.CV2023
Cross-Modal Reasoning with Event Correlation for Video Question Answering
Chengxiang Yin, Zhengping Che, Kun Wu +3
Video Question Answering (VideoQA) is a very attractive and challenging research direction aiming to understand complex semantics of heterogeneous data from two domains, i.e., the…