1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2021
Approximation Methods for Partially Observed Markov Decision Processes (POMDPs)
Caleb M. Bowyer
POMDPs are useful models for systems where the true underlying state is not known completely to an outside observer; the outside observer incompletely knows the true state of the s…
cs.LG2021★ 1 cited
Improving Generalization in Mountain Car Through the Partitioned Parameterized Policy Approach via Quasi-Stochastic Gradient Descent
Caleb M. Bowyer
The reinforcement learning problem of finding a control policy that minimizes the minimum time objective for the Mountain Car environment is considered. Particularly, a class of pa…