18 citations · 59 across the 6 of their papers we have counts for
3 papers · 1 filter
Smooth markets: A basic mechanism for organizing gradient-based learners
David Balduzzi, Wojciech M Czarnecki, Thomas W Anthony +5
With the success of modern machine learning, it is becoming increasingly important to understand and control how learning algorithms interact. Unfortunately, negative results from…
OpenSpiel: A Framework for Reinforcement Learning in Games
Marc Lanctot, Edward Lockhart, Jean-Baptiste Lespiau +24
OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games. OpenSpiel supports n-player (single- and multi…
Policy Gradient Search: Online Planning and Expert Iteration without Search Trees
Thomas Anthony, Robert Nishihara, Philipp Moritz +2
Monte Carlo Tree Search (MCTS) algorithms perform simulation-based search to improve policies online. During search, the simulation policy is adapted to explore the most promising…