1 citations · 1 across the 2 of their papers we have counts for
3 papers · 1 filter
Domain Generalization for Robust Model-Based Offline Reinforcement Learning
Alan Clark, Shoaib Ahmed Siddiqui, Robert Kirk +3
Existing offline reinforcement learning (RL) algorithms typically assume that training data is either: 1) generated by a known policy, or 2) of entirely unknown origin. We consider…
Graph Backup: Data Efficient Backup Exploiting Markovian Transitions
Zhengyao Jiang, Tianjun Zhang, Robert Kirk +2
The successes of deep Reinforcement Learning (RL) are limited to settings where we have a large stream of online experiences, but applying RL in the data-efficient setting with lim…
Insights From the NeurIPS 2021 NetHack Challenge
Eric Hambro, Sharada Mohanty, Dmitrii Babaev +26
In this report, we summarize the takeaways from the first NeurIPS 2021 NetHack Challenge. Participants were tasked with developing a program or agent that can win (i.e., 'ascend' i…