1 paper
Ritesh Goenka, Eashan Gupta, Sushil Khyalia +3
Policy Iteration (PI) is a widely used family of algorithms to compute optimal policies for Markov Decision Problems (MDPs). We derive upper bounds on the running time of PI on Det…