Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
High-fidelity and Lip-synced Talking Face Synthesis via Landmark-based Diffusion Model
Weizhi Zhong, Junfan Lin, Peixin Chen +2
Audio-driven talking face video generation has attracted increasing attention due to its huge industrial potential. Some previous methods focus on learning a direct mapping from au…
cs.CV2024
Visual Tuning
Bruce X. B. Yu, Jianlong Chang, Haixin Wang +9
Fine-tuning visual models has been widely shown promising performance on many downstream visual tasks. With the surprising development of pre-trained visual foundation models, visu…