1 paper
Ondrej Kubicek, Neil Burch, Viliam Lisy
Search in test time is often used to improve the performance of reinforcement learning algorithms. Performing theoretically sound search in fully adversarial two-player games with…