1 paper
Sheng Shi, Xuyang Cao, Jun Zhao +1
In audio-driven video generation, creating Mandarin videos presents significant challenges. Collecting comprehensive Mandarin datasets is difficult, and the complex lip movements i…