activity
20212025
most citedFrom Explicit Communication to Tacit Cooperation:A Novel Paradigm for Cooperative MARL

4 citations · 20 across the 21 of their papers we have counts for

collaborators
Showing 2021Show all

6 papers · 1 filter

cs.AI2021

Cooperative Multi-Agent Reinforcement Learning with Hypergraph Convolution

Yunpeng Bai, Chen Gong, Bin Zhang +3

Recent years have witnessed the great success of multi-agent systems (MAS). Value decomposition, which decomposes joint action values into individual action values, has been an imp…

cs.MA2021★ 2 cited

HAVEN: Hierarchical Cooperative Multi-Agent Reinforcement Learning with Dual Coordination Mechanism

Zhiwei Xu, Yunpeng Bai, Bin Zhang +2

Recently, some challenging tasks in multi-agent systems have been solved by some hierarchical reinforcement learning methods. Inspired by the intra-level and inter-level coordinati…

cs.LG2021

The -Divergence Reinforcement Learning Framework

Chen Gong, Qiang He, Yunpeng Bai +6

The framework of deep reinforcement learning (DRL) provides a powerful and widely applicable mathematical formalization for sequential decision-making. This paper present a novel D…

cs.MA2021★ 1 cited

MMD-MIX: Value Function Factorisation with Maximum Mean Discrepancy for Cooperative Multi-Agent Reinforcement Learning

Zhiwei Xu, Dapeng Li, Yunpeng Bai +1

In the real world, many tasks require multiple agents to cooperate with each other under the condition of local observations. To solve such problems, many multi-agent reinforcement…

cs.MA2021★ 1 cited

SIDE: State Inference for Partially Observable Cooperative Multi-Agent Reinforcement Learning

Zhiwei Xu, Yunpeng Bai, Dapeng Li +2

As one of the solutions to the decentralized partially observable Markov decision process (Dec-POMDP) problems, the value decomposition method has achieved significant results rece…

cs.MA2021

Learning to Coordinate via Multiple Graph Neural Networks

Zhiwei Xu, Bin Zhang, Yunpeng Bai +2

The collaboration between agents has gradually become an important topic in multi-agent systems. The key is how to efficiently solve the credit assignment problems. This paper intr…