4 papers · 1 filter
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
Weifeng Chen, Jiacheng Zhang, Jie Wu +3
The rapid development of diffusion models has triggered diverse applications. Identity-preserving text-to-image generation (ID-T2I) particularly has received significant attention…
MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation
Weimin Wang, Jiawei Liu, Zhijie Lin +9
The growing demand for high-fidelity video generation from textual descriptions has catalyzed significant research in this field. In this work, we introduce MagicVideo-V2 that inte…
UGC: Unified GAN Compression for Efficient Image-to-Image Translation
Yuxi Ren, Jie Wu, Peng Zhang +6
Recent years have witnessed the prevailing progress of Generative Adversarial Networks (GANs) in image-to-image translation. However, the success of these GAN models hinges on pond…
DiffusionEngine: Diffusion Model is Scalable Data Engine for Object Detection
Manlin Zhang, Jie Wu, Yuxi Ren +7
Data is the cornerstone of deep learning. This paper reveals that the recently developed Diffusion Model is a scalable data engine for object detection. Existing methods for scalin…