6 citations · 6 across the 2 of their papers we have counts for
1 paper · 1 filter
Aaron Sidford, Mengdi Wang, Xian Wu +2
In this paper we consider the problem of computing an ε-optimal policy of a discounted Markov Decision Process (DMDP) provided we can only access its transition function through…