1 citations · 1 across the 1 of their papers we have counts for
1 paper
Chao Qu, Hui Li, Chang Liu +6
We propose a \emph{collaborative} multi-agent reinforcement learning algorithm named variational policy propagation (VPP) to learn a \emph{joint} policy through the interactions ov…