3 papers
cs.GT2024
Asymptotic Extinction in Large Coordination Games
Desmond Chan, Bart De Keijzer, Tobias Galla +2
We study the exploration-exploitation trade-off for large multiplayer coordination games where players strategise via Q-Learning, a common learning framework in multi-agent reinfor…
cs.LG2023
AlberDICE: Addressing Out-Of-Distribution Joint Actions in Offline Multi-Agent RL via Alternating Stationary Distribution Correction Estimation
Daiki E. Matsunaga, Jongmin Lee, Jaeseok Yoon +3
One of the main challenges in offline Reinforcement Learning (RL) is the distribution shift that arises from the learned policy deviating from the data collection policy. This is o…
cs.GT2016
On the commitment value and commitment optimal strategies in bimatrix games
Stefanos Leonardos, Costis Melolidakis
Given a bimatrix game, the associated leadership or commitment games are defined as the games at which one player, the leader, commits to a (possibly mixed) strategy and the other…