3 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.CV2024
FreeCompose: Generic Zero-Shot Image Composition with Diffusion Prior
Zhekai Chen, Wen Wang, Zhen Yang +3
We offer a novel approach to image composition, which integrates multiple input images into a single, coherent image. Rather than concentrating on specific use cases such as appear…
cs.CV2024★ 3 cited
Image Textualization: An Automatic Framework for Creating Accurate and Detailed Image Descriptions
Renjie Pi, Jianshu Zhang, Jipeng Zhang +3
Image description datasets play a crucial role in the advancement of various applications such as image understanding, text-to-image generation, and text-image retrieval. Currently…
cs.CV2023★ 2 cited
AutoStory: Generating Diverse Storytelling Images with Minimal Human Effort
Wen Wang, Canyu Zhao, Hao Chen +3
Story visualization aims to generate a series of images that match the story described in texts, and it requires the generated images to satisfy high quality, alignment with the te…