4 papers
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
Rui Chen, Zehuan Wu, Yichen Liu +4
The creation of diverse and realistic driving scenarios has become essential to enhance perception and planning capabilities of the autonomous driving system. However, generating l…
Cross-Block Fine-Grained Semantic Cascade for Skeleton-Based Sports Action Recognition
Zhendong Liu, Haifeng Xia, Tong Guo +3
Human action video recognition has recently attracted more attention in applications such as video security and sports posture correction. Popular solutions, including graph convol…
Embedded Representation Learning Network for Animating Styled Video Portrait
Tianyong Wang, Xiangyu Liang, Wangguandong Zheng +3
The talking head generation recently attracted considerable attention due to its widespread application prospects, especially for digital avatars and 3D animation design. Inspired…
CSTalk: Correlation Supervised Speech-driven 3D Emotional Facial Animation Generation
Xiangyu Liang, Wenlin Zhuang, Tianyong Wang +4
Speech-driven 3D facial animation technology has been developed for years, but its practical application still lacks expectations. The main challenges lie in data limitations, lip…