2 papers
cs.CV2025
YingVideo-MV: Music-Driven Multi-Stage Video Generation
Jiahui Chen, Weida Wang, Runhua Shi +3
While diffusion model for audio-driven avatar video generation have achieved notable process in synthesizing long sequences with natural audio-visual synchronization and identity c…
cs.CV2025
Audio-driven Gesture Generation via Deviation Feature in the Latent Space
Jiahui Chen, Yang Huan, Runhua Shi +4
Gestures are essential for enhancing co-speech communication, offering visual emphasis and complementing verbal interactions. While prior work has concentrated on point-level motio…