1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 1 cited
Optimistic Policy Iteration for MDPs with Acyclic Transient State Structure
Joseph Lubars, Anna Winnicki, Michael Livesay +1
We consider Markov Decision Processes (MDPs) in which every stationary policy induces the same graph structure for the underlying Markov chain and further, the graph has the follow…
math.OC2019
On Privatizing Equilibrium Computation in Aggregate Games over Networks
Shripad Gade, Anna Winnicki, Subhonmesh Bose
We propose a distributed algorithm to compute an equilibrium in aggregate games where players communicate over a fixed undirected network. Our algorithm exploits correlated perturb…