5 citations · 14 across the 3 of their papers we have counts for
4 papers
Multilevel Hierarchical Network with Multiscale Sampling for Video Question Answering
Min Peng, Chongyang Wang, Yuan Gao +2
Video question answering (VideoQA) is challenging given its multimodal combination of visual understanding and natural language processing. While most existing approaches ignore th…
Temporal Pyramid Transformer with Multimodal Interaction for Video Question Answering
Min Peng, Chongyang Wang, Yuan Gao +2
Video question answering (VideoQA) is challenging given its multimodal combination of visual understanding and natural language understanding. While existing approaches seldom leve…
Leveraging Activity Recognition to Enable Protective Behavior Detection in Continuous Data
Chongyang Wang, Yuan Gao, Akhil Mathur +3
Protective behavior exhibited by people with chronic pain (CP) during physical activities is the key to understanding their physical and emotional states. Existing automatic protec…
Recognizing Micro-Expression in Video Clip with Adaptive Key-Frame Mining
Min Peng, Chongyang Wang, Yuan Gao +4
As a spontaneous expression of emotion on face, micro-expression reveals the underlying emotion that cannot be controlled by human. In micro-expression, facial movement is transien…