2 papers
cs.CV2021
Relation-aware Hierarchical Attention Framework for Video Question Answering
Fangtao Li, Ting Bai, Chenyu Cao +3
Video Question Answering (VideoQA) is a challenging video understanding task since it requires a deep understanding of both question and video. Previous studies mainly focus on ext…
cs.CV2020
Frame Aggregation and Multi-Modal Fusion Framework for Video-Based Person Recognition
Fangtao Li, Wenzhe Wang, Zihe Liu +3
Video-based person recognition is challenging due to persons being blocked and blurred, and the variation of shooting angle. Previous research always focused on person recognition…