50 citations · 97 across the 5 of their papers we have counts for
5 papers
Spatial Language Representation with Multi-Level Geocoding
Sayali Kulkarni, Shailee Jain, Mohammad Javad Hosseini +3
We present a multi-level geocoding model (MLG) that learns to associate texts to geographic locations. The Earth's surface is represented using space-filling curves that decompose…
BabyWalk: Going Farther in Vision-and-Language Navigation by Taking Baby Steps
Wang Zhu, Hexiang Hu, Jiacheng Chen +4
Learning to follow instructions is of fundamental importance to autonomous agents for vision-and-language navigation (VLN). In this paper, we study how an agent can navigate long p…
Retouchdown: Adding Touchdown to StreetLearn as a Shareable Resource for Language Grounding Tasks in Street View
Harsh Mehta, Yoav Artzi, Jason Baldridge +2
The Touchdown dataset (Chen et al., 2019) provides instructions by human annotators for navigation through New York City streets and for resolving spatial descriptions at a given l…
VALAN: Vision and Language Agent Navigation
Larry Lansing, Vihan Jain, Harsh Mehta +2
VALAN is a lightweight and scalable software framework for deep reinforcement learning based on the SEED RL architecture. The framework facilitates the development and evaluation o…
RecSim: A Configurable Simulation Platform for Recommender Systems
Eugene Ie, Chih-wei Hsu, Martin Mladenov +5
We propose RecSim, a configurable platform for authoring simulation environments for recommender systems (RSs) that naturally supports sequential interaction with users. RecSim all…