1 paper
Songyang Han, Sanbao Su, Sihong He +4
Various methods for Multi-Agent Reinforcement Learning (MARL) have been developed with the assumption that agents' policies are based on accurate state information. However, polici…