2 papers
cs.CV2026
VINS-120K: Ultra High-Resolution Image Editing with A Large-Scale Dataset
Zhizhou Chen, Shanyan Guan, Zhanxin Gao +6
Directly editing ultra-high-resolution (UHR) images is valuable but underexplored, primarily due to the lack of high-quality data and the challenge in modeling high-frequency textu…
cs.CV2026
CoDi: Subject-Consistent and Pose-Diverse Text-to-Image Generation
Zhanxin Gao, Beier Zhu, Liang Yao +2
Subject-consistent generation (SCG)-aiming to maintain a consistent subject identity across diverse scenes-remains a challenge for text-to-image (T2I) models. Existing training-fre…