5 citations · 5 across the 2 of their papers we have counts for
2 papers
eess.SY2020
Parameter Critic: a Model Free Variance Reduction Method Through Imperishable Samples
Juan Cervino, Harshat Kumar, Alejandro Ribeiro
We consider the problem of finding a policy that maximizes an expected reward throughout the trajectory of an agent that interacts with an unknown environment. Frequently denoted R…
cs.LG2020★ 5 cited
Zeroth-order Deterministic Policy Gradient
Harshat Kumar, Dionysios S. Kalogerias, George J. Pappas +1
Deterministic Policy Gradient (DPG) removes a level of randomness from standard randomized-action Policy Gradient (PG), and demonstrates substantial empirical success for tackling…