1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Istvan Szita, Andras Lorincz
In this paper we propose an algorithm for polynomial-time reinforcement learning in factored Markov decision processes (FMDPs). The factored optimistic initial model (FOIM) algorit…