193 citations · 752 across the 18 of their papers we have counts for
Showing 2021Show all
2 papers · 1 filter
cs.LG2021★ 12 cited
Synthetic Returns for Long-Term Credit Assignment
David Raposo, Sam Ritter, Adam Santoro +5
Since the earliest days of reinforcement learning, the workhorse method for assigning credit to actions over time has been temporal-difference (TD) learning, which propagates credi…
cs.LG2021
Alchemy: A benchmark and analysis toolkit for meta-reinforcement learning agents
Jane X. Wang, Michael King, Nicolas Porcel +14
There has been rapidly growing interest in meta-learning as a method for increasing the flexibility and sample efficiency of reinforcement learning. One problem in this area of res…