9 papers
OmniScaleSR: Unleashing Scale-Controlled Diffusion Prior for Faithful and Realistic Arbitrary-Scale Image Super-Resolution
Xinning Chai, Zhengxue Cheng, Yuhong Zhang +5
Arbitrary-scale super-resolution (ASSR) overcomes the limitation of traditional super-resolution (SR) methods that operate only at fixed scales (e.g., 4x), enabling a single model…
FreeInsert: Personalized Object Insertion with Geometric and Style Control
Yuhong Zhang, Han Wang, Yiwen Wang +2
Text-to-image diffusion models have made significant progress in image generation, allowing for effortless customized generation. However, existing image editing methods still face…
Semantic and Temporal Integration in Latent Diffusion Space for High-Fidelity Video Super-Resolution
Yiwen Wang, Xinning Chai, Yuhong Zhang +4
Recent advancements in video super-resolution (VSR) models have demonstrated impressive results in enhancing low-resolution videos. However, due to limitations in adequately contro…
AnimeColor: Reference-based Animation Colorization with Diffusion Transformers
Yuhong Zhang, Liyao Wang, Han Wang +4
Animation colorization plays a vital role in animation production, yet existing methods struggle to achieve color accuracy and temporal consistency. To address these challenges, we…
NOVA3D: Normal Aligned Video Diffusion Model for Single Image to 3D Generation
Yuxiao Yang, Peihao Li, Yuhong Zhang +5
3D AI-generated content (AIGC) has made it increasingly accessible for anyone to become a 3D content creator. While recent methods leverage Score Distillation Sampling to distill 3…
SSP-IR: Semantic and Structure Priors for Diffusion-based Realistic Image Restoration
Yuhong Zhang, Hengsheng Zhang, Zhengxue Cheng +3
Realistic image restoration is a crucial task in computer vision, and diffusion-based models for image restoration have garnered significant attention due to their ability to produ…