1 paper
Eugene A. Feinberg, Gaojin He
This note provides upper bounds on the number of operations required to compute by value iterations a nearly optimal policy for an infinite-horizon discounted Markov decision proce…