Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Not all tokens contribute equally to diffusion learning
Guoqing Zhang, Lu Shi, Wanru Xu +4
With the rapid development of conditional diffusion models, significant progress has been made in text-to-video generation. However, we observe that these models often neglect sema…
cs.CV2025
Condition Weaving Meets Expert Modulation: Towards Universal and Controllable Image Generation
Guoqing Zhang, Xingtong Ge, Lu Shi +5
The image-to-image generation task aims to produce controllable images by leveraging conditional inputs and prompt instructions. However, existing methods often train separate cont…
cs.CV2024
Patch Spatio-Temporal Relation Prediction for Video Anomaly Detection
Hao Shen, Lu Shi, Wanru Xu +3
Video Anomaly Detection (VAD), aiming to identify abnormalities within a specific context and timeframe, is crucial for intelligent Video Surveillance Systems. While recent deep le…