6 citations · 15 across the 3 of their papers we have counts for
1 paper · 1 filter
Kefan Dong, Jian Peng, Yining Wang +1
In this paper, we consider the problem of online learning of Markov decision processes (MDPs) with very large state spaces. Under the assumptions of realizable function approximati…