1 paper
Yuki Usui, Masahiko Ueda
We investigate the repeated prisoner's dilemma game where both players alternately use reinforcement learning to obtain their optimal memory-one strategies. We theoretically solve…