72 citations · 88 across the 6 of their papers we have counts for
9 papers
Learning to Model Multimodal Semantic Alignment for Story Visualization
Bowen Li, Thomas Lukasiewicz
Story visualization aims to generate a sequence of images to narrate each sentence in a multi-sentence story, where the images should be realistic and keep global consistency acros…
Word-Level Fine-Grained Story Visualization
Bowen Li, Thomas Lukasiewicz
Story visualization aims to generate a sequence of images to narrate each sentence in a multi-sentence story with a global consistency across dynamic scenes and characters. Current…
Lightweight Long-Range Generative Adversarial Networks
Bowen Li, Thomas Lukasiewicz
In this paper, we introduce novel lightweight generative adversarial networks, which can effectively capture long-range dependencies in the image generation process, and produce hi…
Unity Perception: Generate Synthetic Data for Computer Vision
Steve Borkman, Adam Crespi, Saurav Dhakad +12
We introduce the Unity Perception package which aims to simplify and accelerate the process of generating synthetic datasets for computer vision tasks by offering an easy-to-use an…
Lightweight Generative Adversarial Networks for Text-Guided Image Manipulation
Bowen Li, Xiaojuan Qi, Philip H. S. Torr +1
We propose a novel lightweight generative adversarial network for efficient image manipulation using natural language descriptions. To achieve this, a new word-level discriminator…
Knowledge Graph Extraction from Videos
Louis Mahon, Eleonora Giunchiglia, Bowen Li +1
Nearly all existing techniques for automated video annotation (or captioning) describe videos using natural language sentences. However, this has several shortcomings: (i) it is ve…