895 citations · 2k across the 44 of their papers we have counts for
11 papers · 2 filters
StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks
Han Zhang, Tao Xu, Hongsheng Li +4
Synthesizing high-quality images from text descriptions is a challenging problem in computer vision and has many practical applications. Samples generated by existing text-to-image…
Pyramid Scene Parsing Network
Hengshuang Zhao, Jianping Shi, Xiaojuan Qi +2
Scene parsing is challenging for unrestricted open vocabulary and diverse scenes. In this paper, we exploit the capability of global context information by different-region-based c…
CRF-CNN: Modeling Structured Information in Human Pose Estimation
Xiao Chu, Wanli Ouyang, Hongsheng Li +1
Deep convolutional neural networks (CNN) have achieved great success. On the other hand, modeling structural information has been proved critical in many vision problems. It is of…
Crafting GBD-Net for Object Detection
Xingyu Zeng, Wanli Ouyang, Junjie Yan +9
The visual cues from multiple support regions of different sizes and resolutions are complementary in classifying a candidate box in object detection. Effective integration of loca…
Fashion Landmark Detection in the Wild
Ziwei Liu, Sijie Yan, Ping Luo +2
Visual fashion analysis has attracted many attentions in the recent years. Previous work represented clothing regions by either bounding boxes or human joints. This work presents f…
LCrowdV: Generating Labeled Videos for Simulation-based Crowd Behavior Learning
Ernest Cheung, Tsan Kwong Wong, Aniket Bera +2
We present a novel procedural framework to generate an arbitrary number of labeled crowd videos (LCrowdV). The resulting crowd video datasets are used to design accurate algorithms…