1 paper
Lipeng Wan, Zeyang Liu, Xingyu Chen +2
Due to the representation limitation of the joint Q value function, multi-agent reinforcement learning methods with linear value decomposition (LVD) or monotonic value decompositio…