activity
20132023
most citedJointly Attentive Spatial-Temporal Pooling Networks for Video-based Person Re-Identification

40 citations · 310 across the 29 of their papers we have counts for

collaborators
Showing 2018 · cs.CVShow all

5 papers · 2 filters

cs.CV2018

Sequential Attention GAN for Interactive Image Editing

Yu Cheng, Zhe Gan, Yitong Li +2

Most existing text-to-image synthesis tasks are static single-turn generation, based on pre-defined textual descriptions of images. To explore more practical and interactive real-l…

cs.CV2018

StoryGAN: A Sequential Conditional GAN for Story Visualization

Yitong Li, Zhe Gan, Yelong Shen +6

We propose a new task, called Story Visualization. Given a multi-sentence paragraph, the story is visualized by generating a sequence of images, one for each sentence. In contrast…

cs.CV2018

Spatial-Temporal Synergic Residual Learning for Video Person Re-Identification

Xinxing Su, Yingtian Zou, Yu Cheng +3

We tackle the problem of person re-identification in video setting in this paper, which has been viewed as a crucial task in many applications. Meanwhile, it is very challenging si…

cs.CV2018

Pedestrian-Synthesis-GAN: Generating Pedestrian Data in Real Scene and Beyond

Xi Ouyang, Yu Cheng, Yifan Jiang +2

State-of-the-art pedestrian detection models have achieved great success in many benchmarks. However, these models require lots of annotation information and the labeling process u…

cs.CV2018

Understanding Humans in Crowded Scenes: Deep Nested Adversarial Learning and A New Benchmark for Multi-Human Parsing

Jian Zhao, Jianshu Li, Yu Cheng +4

Despite the noticeable progress in perceptual tasks like detection, instance segmentation and human parsing, computers still perform unsatisfactorily on visually understanding huma…