activity
20182022
most citedLearning to Utilize Shaping Rewards: A New Approach of Reward Shaping

94 citations · 140 across the 8 of their papers we have counts for

collaborators
Showing cs.AIShow all

6 papers · 1 filter

cs.AI20223 cited

Revisiting QMIX: Discriminative Credit Assignment by Gradient Entropy Regularization

Jian Zhao, Yue Zhang, Xunhan Hu +5

In cooperative multi-agent systems, agents jointly take actions and receive a team reward instead of individual rewards. In the absence of individual reward signals, credit assignm…

cs.AI202110 cited

Cooperative Multi-Agent Transfer Learning with Level-Adaptive Credit Assignment

Tianze Zhou, Fubiao Zhang, Kun Shao +10

Extending transfer learning to cooperative multi-agent reinforcement learning (MARL) has recently received much attention. In contrast to the single-agent setting, the coordination…

cs.AI2020

KoGuN: Accelerating Deep Reinforcement Learning via Integrating Human Suboptimal Knowledge

Peng Zhang, Jianye Hao, Weixun Wang +4

Reinforcement learning agents usually learn from scratch, which requires a large number of interactions with the environment. This is quite different from the learning process of h…

cs.AI201920 cited

Multi-Agent Game Abstraction via Graph Attention Neural Network

Yong Liu, Weixun Wang, Yujing Hu +3

In large-scale multi-agent systems, the large number of agents and complex game relationship cause great difficulty for policy learning. Therefore, simplifying the learning process…

cs.AI2019

From Few to More: Large-scale Dynamic Multiagent Curriculum Learning

Weixun Wang, Tianpei Yang, Yong Liu +6

A lot of efforts have been devoted to investigating how agents can learn effectively and achieve coordination in multiagent systems. However, it is still challenging in large-scale…

cs.AI2018

Towards Cooperation in Sequential Prisoner's Dilemmas: a Deep Multiagent Reinforcement Learning Approach

Weixun Wang, Jianye Hao, Yixi Wang +1

The Iterated Prisoner's Dilemma has guided research on social dilemmas for decades. However, it distinguishes between only two atomic actions: cooperate and defect. In real-world p…