From the 1 of 15 linked papers with an AI index.
5 papers · 1 filter
Dynamic Resource Allocation for Ensemble Determinization MCTS
Jakub Kowalski, Adam CiÄżkowski, Artur KrzyżyÅski +1
The paper introduces two dynamic resource allocation methods for Ensemble Determinization Monte Carlo Tree Search, adjusting the number of determinization trees and the simulation…
StratFormer: Adaptive Opponent Modeling and Exploitation in Imperfect-Information Games
Andy Caen, Mark H. M. Winands, Dennis J. N. J. Soemers
We present StratFormer, a transformer-based meta-agent that learns to simultaneously model and exploit opponents in imperfect-information games through a two-phase curriculum. The…
A Research Agenda for Usability and Generalisation in Reinforcement Learning
Dennis J. N. J. Soemers, Spyridon Samothrakis, Kurt Driessens +1
It is common practice in reinforcement learning (RL) research to train and deploy agents in bespoke simulators, typically implemented by engineers directly in general-purpose progr…
Generalized Proof-Number Monte-Carlo Tree Search
Jakub Kowalski, Dennis J. N. J. Soemers, Szymon Kosakowski +1
This paper presents Generalized Proof-Number Monte-Carlo Tree Search: a generalization of recently proposed combinations of Proof-Number Search (PNS) with Monte-Carlo Tree Search (…
Towards Explaining Monte-Carlo Tree Search by Using Its Enhancements
Jakub Kowalski, Mark H. M. Winands, Maksymilian WiÅniewski +2
Typically, research on Explainable Artificial Intelligence (XAI) focuses on black-box models within the context of a general policy in a known, specific domain. This paper advocate…