1 paper · 1 filter
Eugene A. Feinberg, Gaojin He
This note provides upper bounds on the number of operations required to compute by value iterations a nearly optimal policy for an infinite-horizon discounted Markov decision proce…