1 paper
Monirul Islam Pavel, Siyi Hu, Muhammad Anwar Masum +3
Real world deployment of multi agent reinforcement learning MARL systems is fundamentally constrained by limited compute memory and inference time. While expert policies achieve hi…