5 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2024
Coordination Failure in Cooperative Offline MARL
Callum Rhys Tilbury, Claude Formanek, Louise Beyers +2
Offline multi-agent reinforcement learning (MARL) leverages static datasets of experience to learn optimal multi-agent control. However, learning from static data presents several…
cs.LG2023★ 5 cited
Revisiting the Gumbel-Softmax in MADDPG
Callum Rhys Tilbury, Filippos Christianos, Stefano V. Albrecht
MADDPG is an algorithm in multi-agent reinforcement learning (MARL) that extends the popular single-agent method, DDPG, to multi-agent scenarios. Importantly, DDPG is an algorithm…