2 papers
cs.CV2020
MGD-GAN: Text-to-Pedestrian generation through Multi-Grained Discrimination
Shengyu Zhang, Donghui Wang, Zhou Zhao +3
In this paper, we investigate the problem of text-to-pedestrian synthesis, which has many potential applications in art, design, and video surveillance. Existing methods for text-t…
cs.CV2018
Attentive Sequence to Sequence Translation for Localizing Clips of Interest by Natural Language Descriptions
Ke Ning, Linchao Zhu, Ming Cai +3
We propose a novel attentive sequence to sequence translator (ASST) for clip localization in videos by natural language descriptions. We make two contributions. First, we propose a…