18 citations · 37 across the 5 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2021★ 2 cited
Average-Reward Reinforcement Learning with Trust Region Methods
Xiaoteng Ma, Xiaohang Tang, Li Xia +2
Most of reinforcement learning algorithms optimize the discounted criterion which is beneficial to accelerate the convergence and reduce the variance of estimates. Although the dis…
cs.LG2021★ 9 cited
Modeling the Interaction between Agents in Cooperative Multi-Agent Reinforcement Learning
Xiaoteng Ma, Yiqin Yang, Chenghao Li +3
Value-based methods of multi-agent reinforcement learning (MARL), especially the value decomposition methods, have been demonstrated on a range of challenging cooperative tasks. Ho…
cs.LG2020
Wasserstein Distance guided Adversarial Imitation Learning with Reward Shape Exploration
Ming Zhang, Yawei Wang, Xiaoteng Ma +4
The generative adversarial imitation learning (GAIL) has provided an adversarial learning framework for imitating expert policy from demonstrations in high-dimensional continuous t…