2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Enhancing Spatial Understanding in Image Generation via Reward Modeling
Zhenyu Tang, Chaoran Feng, Yufan Deng +5
Recent progress in text-to-image generation has greatly advanced visual fidelity and creativity, but it has also imposed higher demands on prompt complexity-particularly in encodin…
cs.CV2024★ 2 cited
MagicVideo-V2: Multi-Stage High-Aesthetic Video Generation
Weimin Wang, Jiawei Liu, Zhijie Lin +9
The growing demand for high-fidelity video generation from textual descriptions has catalyzed significant research in this field. In this work, we introduce MagicVideo-V2 that inte…
cs.CV2023
DreamTuner: Single Image is Enough for Subject-Driven Generation
Miao Hua, Jiawei Liu, Fei Ding +3
Diffusion-based models have demonstrated impressive capabilities for text-to-image generation and are expected for personalized applications of subject-driven generation, which req…