collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing

Jianzong Wu, Hao Lian, Jiongfan Yang +12

Developing unified video generation and editing models capable of interpreting interleaved multimodal inputs is a promising yet challenging frontier field. Existing unified framewo…

cs.CV2026

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers

Jingyuan Zhu, Biaolong Chen, Le Zhang +3

Efficiently aligning large-scale video diffusion models with human intent requires a scalable and trajectory-aware pathway that bridges the inherent discrepancy between training no…

cs.CV2026

ProFashion: Prototype-guided Fashion Video Generation with Multiple Reference Images

Xianghao Kong, Qiaosong Qi, Yuanbin Wang +3

Fashion video generation aims to synthesize temporally consistent videos from reference images of a designated character. Despite significant progress, existing diffusion-based met…

cs.CV2024

FuseAnyPart: Diffusion-Driven Facial Parts Swapping via Multiple Reference Images

Zheng Yu, Yaohua Wang, Siying Cui +3

Facial parts swapping aims to selectively transfer regions of interest from the source image onto the target image while maintaining the rest of the target image unchanged. Most st…

cs.CV2024

RealisDance: Equip controllable character animation with realistic hands

Jingkai Zhou, Benzhi Wang, Weihua Chen +6

Controllable character animation is an emerging task that generates character videos controlled by pose sequences from given character images. Although character consistency has ma…