1 paper · 1 filter
Dan Schmidt, Nick Moran, Jonathan S. Rosenfeld +2
The AlphaZero algorithm for the learning of strategy games via self-play, which has produced superhuman ability in the games of Go, chess, and shogi, uses a quantitative reward fun…