Showing cs.LGShow all
2 papers · 1 filter
cs.LG2019
Self-Play Learning Without a Reward Metric
Dan Schmidt, Nick Moran, Jonathan S. Rosenfeld +2
The AlphaZero algorithm for the learning of strategy games via self-play, which has produced superhuman ability in the games of Go, chess, and shogi, uses a quantitative reward fun…
cs.LG2018
Monotone Learning with Rectified Wire Networks
Veit Elser, Dan Schmidt, Jonathan Yedidia
We introduce a new neural network model, together with a tractable and monotone online learning algorithm. Our model describes feed-forward networks for classification, with one ou…