3 papers
cs.CV2018
Speaker-Follower Models for Vision-and-Language Navigation
Daniel Fried, Ronghang Hu, Volkan Cirik +7
Navigation guided by natural language instructions presents a challenging reasoning problem for instruction followers. Natural language instructions typically identify only a few h…
cs.CL2018
Visual Referring Expression Recognition: What Do Systems Actually Learn?
Volkan Cirik, Louis-Philippe Morency, Taylor Berg-Kirkpatrick
We present an empirical analysis of the state-of-the-art systems for referring expression recognition -- the task of identifying the object in an image referred to by a natural lan…
cs.CV2018
Using Syntax to Ground Referring Expressions in Natural Images
Volkan Cirik, Taylor Berg-Kirkpatrick, Louis-Philippe Morency
We introduce GroundNet, a neural network for referring expression recognition -- the task of localizing (or grounding) in an image the object referred to by a natural language expr…