4 papers
MultiMotion: Multi Subject Video Motion Transfer via Video Diffusion Transformer
Penghui Liu, Jiangshan Wang, Yutong Shen +3
Multi-object video motion transfer poses significant challenges for Diffusion Transformer (DiT) architectures due to inherent motion entanglement and lack of object-level control.…
Unified Long Video Inpainting and Outpainting via Overlapping High-Order Co-Denoising
Shuangquan Lyu, Steven Mao, Yue Ma
Generating long videos remains a fundamental challenge, and achieving high controllability in video inpainting and outpainting is particularly demanding. To address both of these c…
Follow-Your-Preference: Towards Preference-Aligned Image Inpainting
Yutao Shen, Junkun Yuan, Toru Aonishi +2
This paper investigates image inpainting with preference alignment. Instead of introducing a novel method, we go back to basics and revisit fundamental problems in achieving such a…
UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts
Zhen Wan, Chenyang Qi, Zhiheng Liu +2
In this paper, we present UniPaint, a unified generative space-time video inpainting framework that enables spatial-temporal inpainting and interpolation. Different from existing m…