10 citations · 38 across the 11 of their papers we have counts for
1 paper · 1 filter
Yuanheng Zhu, Dongbin Zhao, Mengchen Zhao +1
In single-agent Markov decision processes, an agent can optimize its policy based on the interaction with environment. In multi-player Markov games (MGs), however, the interaction…