4 papers
Talking Head Generation via AU-Guided Landmark Prediction
Shao-Yu Chang, Jingyi Xu, Hieu Le +1
We propose a two-stage framework for audio-driven talking head generation with fine-grained expression control via facial Action Units (AUs). Unlike prior methods relying on emotio…
ACDG-VTON: Accurate and Contained Diffusion Generation for Virtual Try-On
Jeffrey Zhang, Kedan Li, Shao-Yu Chang +1
Virtual Try-on (VTON) involves generating images of a person wearing selected garments. Diffusion-based methods, in particular, can create high-quality images, but they struggle to…
Preserving Image Properties Through Initializations in Diffusion Models
Jeffrey Zhang, Shao-Yu Chang, Kedan Li +1
Retail photography imposes specific requirements on images. For instance, images may need uniform background colors, consistent model poses, centered products, and consistent light…
DiffusionAtlas: High-Fidelity Consistent Diffusion Video Editing
Shao-Yu Chang, Hwann-Tzong Chen, Tyng-Luh Liu
We present a diffusion-based video editing framework, namely DiffusionAtlas, which can achieve both frame consistency and high fidelity in editing video object appearance. Despite…