9 citations · 9 across the 1 of their papers we have counts for
1 paper
Gongbo Zhang, Yijie Peng, Yilong Xu
We consider the popular tree-based search strategy within the framework of reinforcement learning, the Monte Carlo Tree Search (MCTS), in the context of finite-horizon Markov decis…