activity
20182022
most citedSMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving

103 citations · 273 across the 23 of their papers we have counts for

collaborators
Showing cs.MAShow all

10 papers · 1 filter

cs.MA202225 cited

GCS: Graph-based Coordination Strategy for Multi-Agent Reinforcement Learning

Jingqing Ruan, Yali Du, Xuantang Xiong +6

Many real-world scenarios involve a team of agents that have to coordinate their policies to achieve a shared goal. Previous studies mainly focus on decentralized control to maximi…

cs.MA20214 cited

A Game-Theoretic Approach to Multi-Agent Trust Region Optimization

Ying Wen, Hui Chen, Yaodong Yang +4

Trust region methods are widely applied in single-agent reinforcement learning problems due to their monotonic performance-improvement guarantee at every iteration. Nonetheless, wh…

cs.MA202125 cited

MALib: A Parallel Framework for Population-based Multi-agent Reinforcement Learning

Ming Zhou, Ziyu Wan, Hanjing Wang +6

Population-based multi-agent reinforcement learning (PB-MARL) refers to the series of methods nested with reinforcement learning (RL) algorithms, which produces a self-generated se…

cs.MA2021

Learning to Win, Lose and Cooperate through Reward Signal Evolution

Rafal Muszynski, Katja Hofmann, Jun Wang

Solving a reinforcement learning problem typically involves correctly prespecifying the reward signal from which the algorithm learns. Here, we approach the problem of reward signa…

cs.MA2021

Learning in Nonzero-Sum Stochastic Games with Potentials

David Mguni, Yutong Wu, Yali Du +6

Multi-agent reinforcement learning (MARL) has become effective in tackling discrete cooperative game scenarios. However, MARL has yet to penetrate settings beyond those modelled by…

cs.MA2020103 cited

SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving

Ming Zhou, Jun Luo, Julian Villella +34

Multi-agent interaction is a fundamental aspect of autonomous driving in the real world. Despite more than a decade of research and development, the problem of how to competently i…