16 citations · 17 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 16 cited
Discovered Policy Optimisation
Chris Lu, Jakub Grudzien Kuba, Alistair Letcher +3
Tremendous progress has been made in reinforcement learning (RL) over the past decade. Most of these advancements came through the continual development of new algorithms, which we…
cs.LG2022★ 1 cited
Understanding Value Decomposition Algorithms in Deep Cooperative Multi-Agent Reinforcement Learning
Zehao Dou, Jakub Grudzien Kuba, Yaodong Yang
Value function decomposition is becoming a popular rule of thumb for scaling up multi-agent reinforcement learning (MARL) in cooperative games. For such a decomposition rule to hol…