activity
20182021
most citedWhat Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study

107 citations · 161 across the 6 of their papers we have counts for

collaborators

8 papers

cs.LG20211 cited

Braxlines: Fast and Interactive Toolkit for RL-driven Behavior Engineering beyond Reward Maximization

Shixiang Shane Gu, Manfred Diaz, Daniel C. Freeman +7

The goal of continuous control is to synthesize desired behaviors. In reinforcement learning (RL)-driven approaches, this is often accomplished through careful task reward engineer…

cs.RO202121 cited

Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation

C. Daniel Freeman, Erik Frey, Anton Raichuk +3

We present Brax, an open source library for rigid body simulation with a focus on performance and parallelism on accelerators, written in JAX. We present results on a suite of task…

cs.LG202118 cited

What Matters for Adversarial Imitation Learning?

Manu Orsini, Anton Raichuk, Léonard Hussenot +7

Adversarial imitation learning has become a popular framework for imitation in continuous control. Over the years, several variations of its components were proposed to enhance the…

cs.LG20217 cited

Hyperparameter Selection for Imitation Learning

Leonard Hussenot, Marcin Andrychowicz, Damien Vincent +11

We address the issue of tuning hyperparameters (HPs) for imitation learning algorithms in the context of continuous-control, when the underlying reward function of the demonstratin…

cs.LG20217 cited

Agent-Centric Representations for Multi-Agent Reinforcement Learning

Wenling Shang, Lasse Espeholt, Anton Raichuk +1

Object-centric representations have recently enabled significant progress in tackling relational reasoning tasks. By building a strong object-centric inductive bias into neural arc…

cs.LG2020107 cited

What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study

Marcin Andrychowicz, Anton Raichuk, Piotr Stańczyk +9

In recent years, on-policy reinforcement learning (RL) has been successfully applied to many different continuous control tasks. While RL algorithms are often conceptually simple,…