1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.LG2021★ 1 cited
Optimistic Policy Iteration for MDPs with Acyclic Transient State Structure
Joseph Lubars, Anna Winnicki, Michael Livesay +1
We consider Markov Decision Processes (MDPs) in which every stationary policy induces the same graph structure for the underlying Markov chain and further, the graph has the follow…
cs.RO2020
Combining Reinforcement Learning with Model Predictive Control for On-Ramp Merging
Joseph Lubars, Harsh Gupta, Sandeep Chinchali +4
We consider the problem of designing an algorithm to allow a car to autonomously merge on to a highway from an on-ramp. Two broad classes of techniques have been proposed to solve…