activity
20192022
most citedProvably Robust Blackbox Optimization for Reinforcement Learning

11 citations · 35 across the 9 of their papers we have counts for

collaborators

15 papers

cs.LG20221 cited

Learning General World Models in a Handful of Reward-Free Deployments

Yingchen Xu, Jack Parker-Holder, Aldo Pacchiano +5

Building generally capable agents is a grand challenge for deep reinforcement learning (RL). To approach this challenge practically, we outline two key desiderata: 1) to facilitate…

cs.LG20221 cited

Insights From the NeurIPS 2021 NetHack Challenge

Eric Hambro, Sharada Mohanty, Dmitrii Babaev +26

In this report, we summarize the takeaways from the first NeurIPS 2021 NetHack Challenge. Participants were tasked with developing a program or agent that can win (i.e., 'ascend' i…

cs.LG20222 cited

On-the-fly Strategy Adaptation for ad-hoc Agent Coordination

Jaleh Zand, Jack Parker-Holder, Stephen J. Roberts

Training agents in cooperative settings offers the promise of AI agents able to interact effectively with humans (and other agents) in the real world. Multi-agent reinforcement lea…

cs.LG20213 cited

Tuning Mixed Input Hyperparameters on the Fly for Efficient Population Based AutoRL

Jack Parker-Holder, Vu Nguyen, Shaan Desai +1

Despite a series of recent successes in reinforcement learning (RL), many RL algorithms remain sensitive to hyperparameters. As such, there has recently been interest in the field…

cs.LG2021

Augmented World Models Facilitate Zero-Shot Dynamics Generalization From a Single Offline Environment

Philip J. Ball, Cong Lu, Jack Parker-Holder +1

Reinforcement learning from large-scale offline datasets provides us with the ability to learn policies without potentially unsafe or impractical exploration. Significant progress…

cs.LG2021

Unlocking Pixels for Reinforcement Learning via Implicit Attention

Krzysztof Marcin Choromanski, Deepali Jain, Wenhao Yu +9

There has recently been significant interest in training reinforcement learning (RL) agents in vision-based environments. This poses many challenges, such as high dimensionality an…