2 papers
cs.LG2025
Learning and Improving Backgammon Strategy
Gregory R. Galperin
A novel approach to learning is presented, combining features of on-line and off-line methods to achieve considerable performance in the task of learning a backgammon value functio…
cs.LG2025
On-line Policy Improvement using Monte-Carlo Search
Gerald Tesauro, Gregory R. Galperin
We present a Monte-Carlo simulation algorithm for real-time policy improvement of an adaptive controller. In the Monte-Carlo simulation, the long-term expected reward of each possi…