1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Ankush Chakrabarty, Devesh K. Jha, Gregery T. Buzzard +2
We develop a method for obtaining safe initial policies for reinforcement learning via approximate dynamic programming (ADP) techniques for uncertain systems evolving with discrete…