8 citations · 8 across the 2 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2020
Analysis of Hyper-Parameters for Small Games: Iterations or Epochs in Self-Play?
Hui Wang, Michael Emmerich, Mike Preuss +1
The landmark achievements of AlphaGo Zero have created great research interest into self-play in reinforcement learning. In self-play, Monte Carlo Tree Search is used to train a de…
cs.LG2019★ 8 cited
Hyper-Parameter Sweep on AlphaZero General
Hui Wang, Michael Emmerich, Mike Preuss +1
Since AlphaGo and AlphaGo Zero have achieved breakground successes in the game of Go, the programs have been generalized to solve other tasks. Subsequently, AlphaZero was developed…