12 papers · 1 filter
FlowPainter: Inpainting Optical Flow via Confidence-Guided Completion
Yuang Meng, Chenyang Wu, Xianshun Liu +7
Existing optical flow methods broadly follow two paradigms: iterative optimization and diffusion-based estimation. Iterative methods, exemplified by RAFT, achieve high accuracy thr…
YOSE: You Only Select Essential Tokens for Efficient DiT-based Video Object Removal
Chenyang Wu, Lina Lei, Fan Li +6
Recent advances in Diffusion Transformer (DiT)-based video generation technologies have shown impressive results for video object removal. However, these methods still suffer from…
Improving Reconstruction of Representation Autoencoder
Siyu Liu, Chujie Qin, Hubery Yin +6
Recent work leverages Vision Foundation Models as image encoders to boost the generative performance of latent diffusion models (LDMs), as their semantic feature distributions are…
FlowConsist: Make Your Flow Consistent with Real Trajectory
Tianyi Zhang, Chengcheng Liu, Jinwei Chen +5
Fast flow models accelerate the iterative sampling process by learning to directly predict ODE path integrals, enabling one-step or few-step generation. However, we argue that curr…
PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching
Zewei Chang, Zheng-Peng Duan, Jianxing Zhang +6
Image retouching aims to enhance visual quality while aligning with users' personalized aesthetic preferences. To address the challenge of balancing controllability and subjectivit…
VTinker: Guided Flow Upsampling and Texture Mapping for High-Resolution Video Frame Interpolation
Chenyang Wu, Jiayi Fu, Chun-Le Guo +2
Due to large pixel movement and high computational cost, estimating the motion of high-resolution frames is challenging. Thus, most flow-based Video Frame Interpolation (VFI) metho…