4 papers
Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning
Yue Deng, Zirui Wang, Yin Zhang
TD() in value-based MARL algorithms or the Temporal Difference critic learning in Actor-Critic-based (AC-based) algorithms synergistically integrate elements from Monte-Carlo s…
Agent Network Protocol Technical White Paper
Gaowei Chang, Eidan Lin, Chengxuan Yuan +4
With the development of large models and autonomous decision-making AI, agents are rapidly becoming the new entities of the internet, following mobile apps. However, existing inter…
SMAC-R1: The Emergence of Intelligence in Decision-Making Tasks
Yue Deng, Weiyu Ma, Yuxin Fan +4
StarCraft Multi-Agent Challenge (SMAC) has been one of the most commonly used experimental environments in multi-agent reinforcement learning (MARL), where the specific task is to…
SMAC-Hard: Enabling Mixed Opponent Strategy Script and Self-play on SMAC
Yue Deng, Yan Yu, Weiyu Ma +4
The availability of challenging simulation environments is pivotal for advancing the field of Multi-Agent Reinforcement Learning (MARL). In cooperative MARL settings, the StarCraft…