27 citations · 27 across the 1 of their papers we have counts for
1 paper
Rishabh Agarwal, Marlos C. Machado, Pablo Samuel Castro +1
Reinforcement learning methods trained on few environments rarely learn policies that generalize to unseen environments. To improve generalization, we incorporate the inherent sequ…