1 paper
Truong-Huy Dinh Nguyen, Wee-Sun Lee, Tze-Yun Leong
We consider the problem of using a heuristic policy to improve the value approximation by the Upper Confidence Bound applied in Trees (UCT) algorithm in non-adversarial settings su…