activity
20152023
most citedHow to Train Your Robot with Deep Reinforcement Learning; Lessons We've Learned

565 citations · 5.2k across the 138 of their papers we have counts for

collaborators
Showing 2020Show all

49 papers · 1 filter

cs.LG202017 cited

Model-Based Visual Planning with Self-Supervised Functional Distances

Stephen Tian, Suraj Nair, Frederik Ebert +4

A generalist robot must be able to complete a variety of tasks in its environment. One appealing way to specify each task is in terms of a goal observation. However, learning goal-…

cs.LG20201 cited

Variable-Shot Adaptation for Online Meta-Learning

Tianhe Yu, Xinyang Geng, Chelsea Finn +1

Few-shot meta-learning methods consider the problem of learning new tasks from a small, fixed number of examples, by meta-learning across static data from a set of previous tasks.…

cs.LG20204 cited

Models, Pixels, and Rewards: Evaluating Design Trade-offs in Visual Model-Based Reinforcement Learning

Mohammad Babaeizadeh, Mohammad Taghi Saffar, Danijar Hafner +4

Model-based reinforcement learning (MBRL) methods have shown strong sample efficiency and performance across a variety of tasks, including when faced with high-dimensional visual o…

cs.LG202022 cited

Emergent Complexity and Zero-shot Transfer via Unsupervised Environment Design

Michael Dennis, Natasha Jaques, Eugene Vinitsky +4

A wide range of reinforcement learning (RL) problems - including robustness, transfer learning, unsupervised RL, and emergent complexity - require specifying a distribution of task…

cs.LG202027 cited

Parrot: Data-Driven Behavioral Priors for Reinforcement Learning

Avi Singh, Huihan Liu, Gaoyue Zhou +3

Reinforcement learning provides a general framework for flexible decision making and control, but requires extensive data collection for each new task that an agent needs to learn.…

cs.LG20203 cited

Continual Learning of Control Primitives: Skill Discovery via Reset-Games

Kelvin Xu, Siddharth Verma, Chelsea Finn +1

Reinforcement learning has the potential to automate the acquisition of behavior in complex settings, but in order for it to be successfully deployed, a number of practical challen…