1 paper
Ke Zhang, DanDan Zhu, Qiuhan Xu +2
Training for multi-agent reinforcement learning(MARL) is a time-consuming process caused by distribution shift of each agent. One drawback is that strategy of each agent in MARL is…