9 papers
Self-Learning Expression Deformations for Data-Efficient Gaussian Avatars
Jiahao Yang, Xiaohang Yang, Qing Wang +3
Modeling dynamic facial expressions using 3D Gaussian representations remains challenging due to their unstructured nature. Conventional Gaussian avatar pipelines require extensive…
HandSCS: Structural Coordinate Space for Animatable Hand Gaussian Splatting
Yilan Dong, Wenqing Wang, Qing Wang +5
Photorealistic and animatable hand avatars are essential for applications such as AR/VR, gaming, and telepresence. Recent 3D Gaussian Splatting (3DGS) based avatar methods enable r…
DanceChat: Large Language Model-Guided Music-to-Dance Generation
Qing Wang, Xiaohang Yang, Yilan Dong +3
Music-to-dance generation aims to synthesize human dance motion conditioned on musical input. Despite recent progress, significant challenges remain due to the semantic gap between…
Robust Photo-Realistic Hand Gesture Generation: from Single View to Multiple View
Qifan Fu, Xu Chen, Muhammad Asad +3
High-fidelity hand gesture generation represents a significant challenge in human-centric generation tasks. Existing methods typically employ a single-view mesh-rendered image prio…
STaR: Seamless Spatial-Temporal Aware Motion Retargeting with Penetration and Consistency Constraints
Xiaohang Yang, Qing Wang, Jiahao Yang +2
Motion retargeting seeks to faithfully replicate the spatio-temporal motion characteristics of a source character onto a target character with a different body shape. Apart from mo…
SuperCap: Multi-resolution Superpixel-based Image Captioning
Henry Senior, Luca Rossi, Gregory Slabaugh +1
It has been a longstanding goal within image captioning to move beyond a dependence on object detection. We investigate using superpixels coupled with Vision Language Models (VLMs)…