Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
JoyVASA: Portrait and Animal Image Animation with Diffusion-Based Audio-Driven Facial Dynamics and Head Motion Generation
Xuyang Cao, Guoxin Wang, Sheng Shi +5
Audio-driven portrait animation has made significant advances with diffusion-based models, improving video quality and lipsync accuracy. However, the increasing complexity of these…
cs.CV2024
JoyHallo: Digital human model for Mandarin
Sheng Shi, Xuyang Cao, Jun Zhao +1
In audio-driven video generation, creating Mandarin videos presents significant challenges. Collecting comprehensive Mandarin datasets is difficult, and the complex lip movements i…