6 citations · 11 across the 2 of their papers we have counts for
2 papers
cs.LG2023★ 6 cited
MAC-PO: Multi-Agent Experience Replay via Collective Priority Optimization
Yongsheng Mei, Hanhan Zhou, Tian Lan +2
Experience replay is crucial for off-policy reinforcement learning (RL) methods. By remembering and reusing the experiences from past different policies, experience replay signific…
cs.LG2023★ 5 cited
ReMIX: Regret Minimization for Monotonic Value Function Factorization in Multiagent Reinforcement Learning
Yongsheng Mei, Hanhan Zhou, Tian Lan
Value function factorization methods have become a dominant approach for cooperative multiagent reinforcement learning under a centralized training and decentralized execution para…