3 papers
cs.CV2025
CoMo: Compositional Motion Customization for Text-to-Video Generation
Youcan Xu, Zhen Wang, Jiaxin Shi +6
While recent text-to-video models excel at generating diverse scenes, they struggle with precise motion control, particularly for complex, multi-subject motions. Although methods f…
cs.CR2025
TarPro: Targeted Protection against Malicious Image Editing
Kaixin Shen, Ruijie Quan, Jiaxu Miao +2
The rapid advancement of image editing techniques has raised concerns about their misuse for generating Not-Safe-for-Work (NSFW) content. This necessitates a targeted protection me…
cs.CV2024
Collaborative Hybrid Propagator for Temporal Misalignment in Audio-Visual Segmentation
Kexin Li, Zongxin Yang, Yi Yang +1
Audio-visual video segmentation (AVVS) aims to generate pixel-level maps of sound-producing objects that accurately align with the corresponding audio. However, existing methods of…