1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Yuanjun Li, Bin Zhang, Hao Chen +3
Value decomposition (VD) methods have achieved remarkable success in cooperative multi-agent reinforcement learning (MARL). However, their reliance on the max operator for temporal…