2 citations · 2 across the 4 of their papers we have counts for
4 papers
AUD-TGN: Advancing Action Unit Detection with Temporal Convolution and GPT-2 in Wild Audiovisual Contexts
Jun Yu, Zerui Zhang, Zhihong Wei +6
Leveraging the synergy of both audio data and visual data is essential for understanding human emotions and behaviors, especially in in-the-wild setting. Traditional methods for in…
Multimodal Fusion Method with Spatiotemporal Sequences and Relationship Learning for Valence-Arousal Estimation
Jun Yu, Gongpeng Zhao, Yongqi Wang +7
This paper presents our approach for the VA (Valence-Arousal) estimation task in the ABAW6 competition. We devised a comprehensive model by preprocessing video frames and audio seg…
Exploring Facial Expression Recognition through Semi-Supervised Pretraining and Temporal Modeling
Jun Yu, Zhihong Wei, Zhongpeng Cai +6
Facial Expression Recognition (FER) plays a crucial role in computer vision and finds extensive applications across various fields. This paper aims to present our approach for the…
RHOBIN Challenge: Reconstruction of Human Object Interaction
Xianghui Xie, Xi Wang, Nikos Athanasiou +17
Modeling the interaction between humans and objects has been an emerging research direction in recent years. Capturing human-object interaction is however a very challenging task d…