1 paper
Maxim Mednikov, Oren Gal
Real-world multi-agent reinforcement learning (MARL) systems must often operate under stale observations, stochastic communication delays, and intermittent packet loss. Policies tr…