Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
HO-Cap: A Capture System and Dataset for 3D Reconstruction and Pose Tracking of Hand-Object Interaction
Jikai Wang, Qifan Zhang, Yu-Wei Chao +3
We introduce a data capture system and a new dataset, HO-Cap, for 3D reconstruction and pose tracking of hands and objects in videos. The system leverages multiple RGBD cameras and…
cs.CV2024
Joint Co-Speech Gesture and Expressive Talking Face Generation using Diffusion with Adapters
Steven Hogue, Chenxu Zhang, Yapeng Tian +1
Recent advances in co-speech gesture and talking head generation have been impressive, yet most methods focus on only one of the two tasks. Those that attempt to generate both ofte…
cs.CV2024
DiffTED: One-shot Audio-driven TED Talk Video Generation with Diffusion-based Co-speech Gestures
Steven Hogue, Chenxu Zhang, Hamza Daruger +2
Audio-driven talking video generation has advanced significantly, but existing methods often depend on video-to-video translation techniques and traditional generative networks lik…