2 papers
cs.CV2024
Real-time One-Step Diffusion-based Expressive Portrait Videos Generation
Hanzhong Guo, Hongwei Yi, Daquan Zhou +3
Latent diffusion models have made great strides in generating expressive portrait videos with accurate lip-sync and natural motion from a single reference image and audio input. Ho…
cs.LG2024
Phased Consistency Models
Fu-Yun Wang, Zhaoyang Huang, Alexander William Bergman +9
Consistency Models (CMs) have made significant progress in accelerating the generation of diffusion models. However, their application to high-resolution, text-conditioned image ge…