3 papers
cs.CV2025
SMRABooth: Subject and Motion Representation Alignment for Customized Video Generation
Xuancheng Xu, Yaning Li, Sisi You +1
Customized video generation aims to produce videos that faithfully preserve the subject's appearance from reference images while maintaining temporally consistent motion from refer…
cs.CV2025
Chain-of-Cooking:Cooking Process Visualization via Bidirectional Chain-of-Thought Guidance
Mengling Xu, Ming Tao, Bing-Kun Bao
Cooking process visualization is a promising task in the intersection of image generation and food analysis, which aims to generate an image for each cooking step of a recipe. Howe…
cs.CV2024
Do We Need to Design Specific Diffusion Models for Different Tasks? Try ONE-PIC
Ming Tao, Bing-Kun Bao, Yaowei Wang +1
Large pretrained diffusion models have demonstrated impressive generation capabilities and have been adapted to various downstream tasks. However, unlike Large Language Models (LLM…