5 papers
Seedance 2.0: Advancing Video Generation for World Complexity
Team Seedance, De Chen, Liyang Chen +168
Seedance 2.0 is a new native multi-modal audio-video generation model, officially released in China in early February 2026. Compared with its predecessors, Seedance 1.0 and 1.5 Pro…
DreamO: A Unified Framework for Image Customization
Chong Mou, Yanze Wu, Wenxu Wu +15
Recently, extensive research on image customization (e.g., identity, subject, style, background, etc.) demonstrates strong customization capabilities in large-scale generative mode…
UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward
Yufeng Cheng, Wenxu Wu, Shaojin Wu +3
Recent advancements in image customization exhibit a wide range of application prospects due to stronger customization capabilities. However, since we humans are more sensitive to…
USO: Unified Style and Subject-Driven Generation via Disentangled and Reward Learning
Shaojin Wu, Mengqi Huang, Yufeng Cheng +5
Existing literature typically treats style-driven and subject-driven generation as two disjoint tasks: the former prioritizes stylistic similarity, whereas the latter insists on su…
Less-to-More Generalization: Unlocking More Controllability by In-Context Generation
Shaojin Wu, Mengqi Huang, Wenxu Wu +3
Although subject-driven generation has been extensively explored in image generation due to its wide applications, it still has challenges in data scalability and subject expansibi…