◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Thomas W. Anthony

8 papers hereh-index 121.5k citations21 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author6

Across the 7 of 8 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.GT2
  • cs.MA2
  • cs.AI1

identity via Semantic Scholar / OpenAlex

activity
20192022
most citedPolicy Gradient Search: Online Planning and Expert Iteration without Search Trees

18 citations · 59 across the 6 of their papers we have counts for

collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2020★ 7 cited

Smooth markets: A basic mechanism for organizing gradient-based learners

David Balduzzi, Wojciech M Czarnecki, Thomas W Anthony +5

With the success of modern machine learning, it is becoming increasingly important to understand and control how learning algorithms interact. Unfortunately, negative results from…

cs.LG2019

OpenSpiel: A Framework for Reinforcement Learning in Games

Marc Lanctot, Edward Lockhart, Jean-Baptiste Lespiau +24

OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games. OpenSpiel supports n-player (single- and multi…

cs.LG2019★ 18 cited

Policy Gradient Search: Online Planning and Expert Iteration without Search Trees

Thomas Anthony, Robert Nishihara, Philipp Moritz +2

Monte Carlo Tree Search (MCTS) algorithms perform simulation-based search to improve policies online. During search, the simulation policy is adapted to explore the most promising…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.