activity
20152023
most citedHow to Train Your Robot with Deep Reinforcement Learning; Lessons We've Learned

565 citations · 5.2k across the 134 of their papers we have counts for

collaborators
Showing 2022Show all

20 papers · 1 filter

cs.LG2022

Data-Driven Offline Decision-Making via Invariant Representation Learning

Han Qi, Yi Su, Aviral Kumar +1

The goal in offline data-driven decision-making is synthesize decisions that optimize a black-box utility function, using a previously-collected static dataset, with no active inte…

cs.LG20222 cited

Offline RL With Realistic Datasets: Heteroskedasticity and Support Constraints

Anikait Singh, Aviral Kumar, Quan Vuong +2

Offline reinforcement learning (RL) learns policies entirely from static datasets, thereby avoiding the challenges associated with online data collection. Practical applications of…

cs.LG20222 cited

Dual Generator Offline Reinforcement Learning

Quan Vuong, Aviral Kumar, Sergey Levine +1

In offline RL, constraining the learned policy to remain close to the data is essential to prevent the policy from outputting out-of-distribution (OOD) actions with erroneously ove…

cs.LG202212 cited

Unpacking Reward Shaping: Understanding the Benefits of Reward Engineering on Sample Complexity

Abhishek Gupta, Aldo Pacchiano, Yuexiang Zhai +2

Reinforcement learning provides an automated framework for learning behaviors from high-level reward specifications, but in practice the choice of reward function can be crucial fo…

cs.LG20223 cited

You Only Live Once: Single-Life Reinforcement Learning

Annie S. Chen, Archit Sharma, Sergey Levine +1

Reinforcement learning algorithms are typically designed to learn a performant policy that can repeatedly and autonomously complete a task, usually starting from scratch. However,…

cs.RO20222 cited

ExAug: Robot-Conditioned Navigation Policies via Geometric Experience Augmentation

Noriaki Hirose, Dhruv Shah, Ajay Sridhar +1

Machine learning techniques rely on large and diverse datasets for generalization. Computer vision, natural language processing, and other applications can often reuse public datas…