416 citations · 840 across the 24 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2023
Emage: Non-Autoregressive Text-to-Image Generation
Zhangyin Feng, Runyi Hu, Liangxin Liu +7
Autoregressive and diffusion models drive the recent breakthroughs on text-to-image generation. Despite their huge success of generating high-realistic images, a common shortcoming…
cs.CV2023
Improved Visual Story Generation with Adaptive Context Modeling
Zhangyin Feng, Yuchen Ren, Xinmiao Yu +4
Diffusion models developed on top of powerful text-to-image generation models like Stable Diffusion achieve remarkable success in visual story generation. However, the best-perform…
cs.CV2019★ 2 cited
Deep Reason: A Strong Baseline for Real-World Visual Reasoning
Chenfei Wu, Yanzhao Zhou, Gen Li +3
This paper presents a strong baseline for real-world visual reasoning (GQA), which achieves 60.93% in GQA 2019 challenge and won the sixth place. GQA is a large dataset with 22M qu…