activity
20172021
most citedTips and Tricks for Visual Question Answering: Learnings from the 2017 Challenge

2 citations · 2 across the 1 of their papers we have counts for

collaborators

8 papers

cs.CV2021

Pathdreamer: A World Model for Indoor Navigation

Jing Yu Koh, Honglak Lee, Yinfei Yang +2

People navigating in unfamiliar buildings take advantage of myriad visual, spatial and semantic cues to efficiently achieve their navigation goals. Towards equipping computational…

cs.CV2019

Chasing Ghosts: Instruction Following as Bayesian State Tracking

Peter Anderson, Ayush Shrivastava, Devi Parikh +2

A visually-grounded navigation instruction can be interpreted as a sequence of expected observations and actions an agent following the correct trajectory would encounter and perfo…

cs.CV2018

nocaps: novel object captioning at scale

Harsh Agrawal, Karan Desai, Yufei Wang +7

Image captioning models have achieved impressive results on datasets containing limited visual concepts and large amounts of paired image-caption training data. However, if these m…

cs.CL2018

Disfluency Detection using Auto-Correlational Neural Networks

Paria Jamshid Lou, Peter Anderson, Mark Johnson

In recent years, the natural language processing community has moved away from task-specific feature engineering, i.e., researchers discovering ad-hoc feature representations for v…

cs.AI2018

On Evaluation of Embodied Navigation Agents

Peter Anderson, Angel Chang, Devendra Singh Chaplot +8

Skillful mobile operation in three-dimensional environments is a primary topic of study in Artificial Intelligence. The past two years have seen a surge of creative work on navigat…

cs.CV2018

Face-Cap: Image Captioning using Facial Expression Analysis

Omid Mohamad Nezami, Mark Dras, Peter Anderson +1

Image captioning is the process of generating a natural language description of an image. Most current image captioning models, however, do not take into account the emotional aspe…