28 citations · 33 across the 2 of their papers we have counts for
4 papers
Temporal Pyramid Transformer with Multimodal Interaction for Video Question Answering
Min Peng, Chongyang Wang, Yuan Gao +2
Video question answering (VideoQA) is challenging given its multimodal combination of visual understanding and natural language understanding. While existing approaches seldom leve…
STA-VPR: Spatio-temporal Alignment for Visual Place Recognition
Feng Lu, Baifan Chen, Xiang-Dong Zhou +1
Recently, the methods based on Convolutional Neural Networks (CNNs) have gained popularity in the field of visual place recognition (VPR). In particular, the features from the midd…
Recognizing Micro-Expression in Video Clip with Adaptive Key-Frame Mining
Min Peng, Chongyang Wang, Yuan Gao +4
As a spontaneous expression of emotion on face, micro-expression reveals the underlying emotion that cannot be controlled by human. In micro-expression, facial movement is transien…
A Novel Apex-Time Network for Cross-Dataset Micro-Expression Recognition
Min Peng, Chongyang Wang, Tao Bi +3
The automatic recognition of micro-expression has been boosted ever since the successful introduction of deep learning approaches. As researchers working on such topics are moving…