33 citations · 70 across the 5 of their papers we have counts for
5 papers
Semantics Disentangling for Text-to-Image Generation
Guojun Yin, Bin Liu, Lu Sheng +3
Synthesizing photo-realistic images from text descriptions is a challenging problem. Previous studies have shown remarkable progresses on visual quality of the generated images. In…
Context and Attribute Grounded Dense Captioning
Guojun Yin, Lu Sheng, Bin Liu +3
Dense captioning aims at simultaneously localizing semantic regions and describing these regions-of-interest (ROIs) with short phrases or sentences in natural language. Previous st…
GS3D: An Efficient 3D Object Detection Framework for Autonomous Driving
Buyu Li, Wanli Ouyang, Lu Sheng +2
We present an efficient 3D object detection framework based on a single RGB image in the scenario of autonomous driving. Our efforts are put on extracting the underlying 3D informa…
Video Generation from Single Semantic Label Map
Junting Pan, Chengyu Wang, Xu Jia +4
This paper proposes the novel task of video generation conditioned on a SINGLE semantic label map, which provides a good balance between flexibility and quality in the generation p…
Unsupervised Bi-directional Flow-based Video Generation from one Snapshot
Lu Sheng, Junting Pan, Jiaming Guo +3
Imagining multiple consecutive frames given one single snapshot is challenging, since it is difficult to simultaneously predict diverse motions from a single image and faithfully g…