69 citations · 402 across the 23 of their papers we have counts for
1 paper · 1 filter
Aaron Sidford, Mengdi Wang, Xian Wu +2
In this paper we consider the problem of computing an ε-optimal policy of a discounted Markov Decision Process (DMDP) provided we can only access its transition function through…