Memory-two strategies forming symmetric mutual reinforcement learning equilibrium in repeated prisoners' dilemma game
arXiv:2108.03258 · doi:10.1016/j.amc.2022.127819
Abstract
We investigate symmetric equilibria of mutual reinforcement learning when both players alternately learn the optimal memory-two strategies against the opponent in the repeated prisoners' dilemma game. We provide a necessary condition for memory-two deterministic strategies to form symmetric equilibria. We then provide three examples of memory-two deterministic strategies which form symmetric mutual reinforcement learning equilibria. We also prove that mutual reinforcement learning equilibria formed by memory-two strategies are also mutual reinforcement learning equilibria when both players use reinforcement learning of memory- strategies with .
26 pages
References in corpus (7)
- Evolutionary dynamics of group interactions on structured populations: A review
- Phase diagrams for three-strategy evolutionary prisoner's dilemma games on regular graphs
- Numerical analysis of a reinforcement learning model with the dynamic aspiration level in the iterated Prisoner's Dilemma
- Combination with anti-tit-for-tat remedies problems of tit-for-tat
- Five rules for friendly rivalry in direct reciprocity
- Memory-two zero-determinant strategies in repeated games
- Symmetric equilibrium of multi-agent reinforcement learning in repeated prisoner's dilemma