2 citations · 2 across the 4 of their papers we have counts for
1 paper · 1 filter
Michael Tryfan Matthews, Anssi Kanervisto, Jakob Foerster +3
Recent work in hierarchical reinforcement learning has shown success in scaling to billions of timesteps when learning over a set of predefined option reward functions. We show tha…