1 citations · 1 across the 1 of their papers we have counts for
1 paper
Feng Zhu, Robert W. Heath, Aritra Mitra
We consider a setting involving N agents, where each agent interacts with an environment modeled as a Markov Decision Process (MDP). The agents' MDPs differ in their reward funct…