3 citations · 6 across the 4 of their papers we have counts for
4 papers
CTDS: Centralized Teacher with Decentralized Student for Multi-Agent Reinforcement Learning
Jian Zhao, Xunhan Hu, Mingyu Yang +3
Due to the partial observability and communication constraints in many multi-agent reinforcement learning (MARL) tasks, centralized training with decentralized execution (CTDE) has…
Revisiting QMIX: Discriminative Credit Assignment by Gradient Entropy Regularization
Jian Zhao, Yue Zhang, Xunhan Hu +5
In cooperative multi-agent systems, agents jointly take actions and receive a team reward instead of individual rewards. In the absence of individual reward signals, credit assignm…
MCMARL: Parameterizing Value Function via Mixture of Categorical Distributions for Multi-Agent Reinforcement Learning
Jian Zhao, Mingyu Yang, Youpeng Zhao +4
In cooperative multi-agent tasks, a team of agents jointly interact with an environment by taking actions, receiving a team reward and observing the next state. During the interact…
An Optimal Resource Allocator of Elastic Training for Deep Learning Jobs on Cloud
Liang Hu, Jiangcheng Zhu, Zirui Zhou +3
Cloud training platforms, such as Amazon Web Services and Huawei Cloud provide users with computational resources to train their deep learning jobs. Elastic training is a service e…