1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Siow Meng Low, Akshat Kumar, Scott Sanner
Recent advances in deep learning have enabled optimization of deep reactive policies (DRPs) for continuous MDP planning by encoding a parametric policy as a deep neural network and…