28 citations · 28 across the 2 of their papers we have counts for
3 papers
cs.CV2023
From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Recognition in Videos
Yin Chen, Jia Li, Shiguang Shan +2
Dynamic facial expression recognition (DFER) in the wild is still hindered by data limitations, e.g., insufficient quantity and diversity of pose, occlusion and illumination, as we…
cs.MM2023★ 28 cited
Exploring Sparse Spatial Relation in Graph Inference for Text-Based VQA
Sheng Zhou, Dan Guo, Jia Li +2
Text-based visual question answering (TextVQA) faces the significant challenge of avoiding redundant relational inference. To be specific, a large number of detected objects and op…
cs.CV2023
Dual-Path Temporal Map Optimization for Make-up Temporal Video Grounding
Jiaxiu Li, Kun Li, Jia Li +3
Make-up temporal video grounding (MTVG) aims to localize the target video segment which is semantically related to a sentence describing a make-up activity, given a long video. Com…