1 paper
Wenhao Zhan, Scott Fujimoto, Zheqing Zhu +3
We study the problem of learning an approximate equilibrium in the offline multi-agent reinforcement learning (MARL) setting. We introduce a structural assumption -- the interactio…