7 papers
MagicFight: Personalized Martial Arts Combat Video Generation
Jiancheng Huang, Mingfu Yan, Songyan Chen +2
Amid the surge in generic text-to-video generation, the field of personalized human video generation has witnessed notable advancements, primarily concentrated on single-person sce…
Contrast-Prior Enhanced Duality for Mask-Free Shadow Removal
Jiyu Wu, Yifan Liu, Jiancheng Huang +2
Existing shadow removal methods often rely on shadow masks, which are challenging to acquire in real-world scenarios. Exploring intrinsic image cues, such as local contrast informa…
Leveraging Contrast Information for Efficient Document Shadow Removal
Yifan Liu, Jiancheng Huang, Na Liu +3
Document shadows are a major obstacle in the digitization process. Due to the dense information in text and patterns covered by shadows, document shadow removal requires specialize…
TARA: Token-Aware LoRA for Composable Personalization in Diffusion Models
Yuqi Peng, Lingtao Zheng, Yufeng Yang +4
Personalized text-to-image generation aims to synthesize novel images of a specific subject or style using only a few reference images. Recent methods based on Low-Rank Adaptation…
DIVE: Taming DINO for Subject-Driven Video Editing
Yi Huang, Wei Xiong, He Zhang +4
Building on the success of diffusion models in image generation and editing, video editing has recently gained substantial attention. However, maintaining temporal consistency and…
Component Adaptive Clustering for Generalized Category Discovery
Mingfu Yan, Jiancheng Huang, Yifan Liu +1
Generalized Category Discovery (GCD) tackles the challenging problem of categorizing unlabeled images into both known and novel classes within a partially labeled dataset, without…