185 citations · 254 across the 5 of their papers we have counts for
1 paper · 2 filters
Bozhidar Vasilev, Tarun Gupta, Bei Peng +1
Policy gradient methods are an attractive approach to multi-agent reinforcement learning problems due to their convergence properties and robustness in partially observable scenari…