2 papers
cs.AI2020
POLY-HOOT: Monte-Carlo Planning in Continuous Space MDPs with Non-Asymptotic Analysis
Weichao Mao, Kaiqing Zhang, Qiaomin Xie +1
Monte-Carlo planning, as exemplified by Monte-Carlo Tree Search (MCTS), has demonstrated remarkable performance in applications with finite spaces. In this paper, we consider Monte…
cs.AI2020
Information State Embedding in Partially Observable Cooperative Multi-Agent Reinforcement Learning
Weichao Mao, Kaiqing Zhang, Erik Miehling +1
Multi-agent reinforcement learning (MARL) under partial observability has long been considered challenging, primarily due to the requirement for each agent to maintain a belief ove…