72 citations · 88 across the 6 of their papers we have counts for
10 papers · 1 filter
ChartGalaxy: A Dataset for Infographic Chart Understanding and Generation
Zhen Li, Duan Li, Yukai Guo +9
Infographic charts are a powerful medium for communicating abstract data by combining visual elements (e.g., charts, images) with textual information. However, their visual and str…
Optimisation-Based Multi-Modal Semantic Image Editing
Bowen Li, Yongxin Yang, Steven McDonagh +3
Image editing affords increased control over the aesthetics and content of generated images. Pre-existing works focus predominantly on text-based instructions to achieve desired im…
Learning to Model Multimodal Semantic Alignment for Story Visualization
Bowen Li, Thomas Lukasiewicz
Story visualization aims to generate a sequence of images to narrate each sentence in a multi-sentence story, where the images should be realistic and keep global consistency acros…
Word-Level Fine-Grained Story Visualization
Bowen Li, Thomas Lukasiewicz
Story visualization aims to generate a sequence of images to narrate each sentence in a multi-sentence story with a global consistency across dynamic scenes and characters. Current…
Lightweight Long-Range Generative Adversarial Networks
Bowen Li, Thomas Lukasiewicz
In this paper, we introduce novel lightweight generative adversarial networks, which can effectively capture long-range dependencies in the image generation process, and produce hi…
Unity Perception: Generate Synthetic Data for Computer Vision
Steve Borkman, Adam Crespi, Saurav Dhakad +12
We introduce the Unity Perception package which aims to simplify and accelerate the process of generating synthetic datasets for computer vision tasks by offering an easy-to-use an…