14 citations · 18 across the 4 of their papers we have counts for
4 papers
X2Video: Adapting Diffusion Models for Multimodal Controllable Neural Video Rendering
Zhitong Huang, Mohan Zhang, Renhan Wang +3
We present X2Video, the first diffusion model for rendering photorealistic videos guided by intrinsic channels including albedo, normal, roughness, metallicity, and irradiance, whi…
LVCD: Reference-based Lineart Video Colorization with Diffusion Models
Zhitong Huang, Mohan Zhang, Jing Liao
We propose the first video diffusion framework for reference-based lineart video colorization. Unlike previous works that rely solely on image generative models to colorize lineart…
Continuous Layout Editing of Single Images with Diffusion Models
Zhiyuan Zhang, Zhitong Huang, Jing Liao
Recent advancements in large-scale text-to-image diffusion models have enabled many applications in image editing. However, none of these methods have been able to edit the layout…
UniColor: A Unified Framework for Multi-Modal Colorization with Transformer
Zhitong Huang, Nanxuan Zhao, Jing Liao
We propose the first unified framework UniColor to support colorization in multiple modalities, including both unconditional and conditional ones, such as stroke, exemplar, text, a…