activity
20172019
most citedObject-driven Text-to-Image Synthesis via Adversarial Training

46 citations · 53 across the 3 of their papers we have counts for

collaborators

10 papers

cs.CL2019

Mapping Natural-language Problems to Formal-language Solutions Using Structured Neural Representations

Kezhen Chen, Qiuyuan Huang, Hamid Palangi +3

Generating formal-language programs represented by relational tuples, such as Lisp programs or mathematical operations, to solve problems stated in natural language is a challengin…

cs.CL2019

REO-Relevance, Extraness, Omission: A Fine-grained Evaluation for Image Captioning

Ming Jiang, Junjie Hu, Qiuyuan Huang +3

Popular metrics used for evaluating image captioning systems, such as BLEU and CIDEr, provide a single score to gauge the system's overall effectiveness. This score is often not in…

cs.CL2019

TIGEr: Text-to-Image Grounding for Image Caption Evaluation

Ming Jiang, Qiuyuan Huang, Lei Zhang +5

This paper presents a new metric called TIGEr for the automatic evaluation of image captioning systems. Popular metrics, such as BLEU and CIDEr, are based solely on text matching b…

cs.CV201946 cited

Object-driven Text-to-Image Synthesis via Adversarial Training

Wenbo Li, Pengchuan Zhang, Lei Zhang +4

In this paper, we propose Object-driven Attentive Generative Adversarial Newtorks (Obj-GANs) that allow object-centered text-to-image synthesis for complex scenes. Following the tw…

cs.CV2018

Reinforced Cross-Modal Matching and Self-Supervised Imitation Learning for Vision-Language Navigation

Xin Wang, Qiuyuan Huang, Asli Celikyilmaz +5

Vision-language navigation (VLN) is the task of navigating an embodied agent to carry out natural language instructions inside real 3D environments. In this paper, we study how to…

cs.CV2018

Hierarchically Structured Reinforcement Learning for Topically Coherent Visual Story Generation

Qiuyuan Huang, Zhe Gan, Asli Celikyilmaz +3

We propose a hierarchically structured reinforcement learning approach to address the challenges of planning for generating coherent multi-sentence stories for the visual storytell…