3 papers
cs.LG2021
No-Press Diplomacy from Scratch
Anton Bakhtin, David Wu, Adam Lerer +1
Prior AI successes in complex games have largely focused on settings with at most hundreds of actions at each decision point. In contrast, Diplomacy is a game with more than 10^20…
cs.AI2021
Off-Belief Learning
Hengyuan Hu, Adam Lerer, Brandon Cui +4
The standard problem setting in Dec-POMDPs is self-play, where the goal is to find a set of policies that play optimally together. Policies learned through self-play may adopt arbi…
cs.LG2019
Accelerating Self-Play Learning in Go
David J. Wu
By introducing several improvements to the AlphaZero process and architecture, we greatly accelerate self-play learning in Go, achieving a 50x reduction in computation over compara…