24 citations · 32 across the 9 of their papers we have counts for
1 paper · 1 filter
Phillip Swazinna, Steffen Udluft, Daniel Hein +1
Offline reinforcement learning (RL) Algorithms are often designed with environments such as MuJoCo in mind, in which the planning horizon is extremely long and no noise exists. We…