39 citations · 77 across the 8 of their papers we have counts for
1 paper · 1 filter
Chen-Yu Wei, Yi-Te Hong, Chi-Jen Lu
We study online reinforcement learning in average-reward stochastic games (SGs). An SG models a two-player zero-sum game in a Markov environment, where state transitions and one-st…