9 citations · 16 across the 2 of their papers we have counts for
3 papers
cs.LG2021★ 9 cited
Modeling the Interaction between Agents in Cooperative Multi-Agent Reinforcement Learning
Xiaoteng Ma, Yiqin Yang, Chenghao Li +3
Value-based methods of multi-agent reinforcement learning (MARL), especially the value decomposition methods, have been demonstrated on a range of challenging cooperative tasks. Ho…
cs.AI2020★ 7 cited
SOAC: The Soft Option Actor-Critic Architecture
Chenghao Li, Xiaoteng Ma, Chongjie Zhang +3
The option framework has shown great promise by automatically extracting temporally-extended sub-tasks from a long-horizon task. Methods have been proposed for concurrently learnin…
cs.LG2020
Wasserstein Distance guided Adversarial Imitation Learning with Reward Shape Exploration
Ming Zhang, Yawei Wang, Xiaoteng Ma +4
The generative adversarial imitation learning (GAIL) has provided an adversarial learning framework for imitating expert policy from demonstrations in high-dimensional continuous t…