1 paper
Yi Xiong, Ningyuan Chen, Xuefeng Gao
When two players are engaged in a repeated game with unknown payoff matrices, they may use single-agent multi-armed bandit algorithms to choose the actions independent of each othe…