1 citations · 1 across the 1 of their papers we have counts for
1 paper
Stephen McAleer, JB Lanier, Kevin Wang +3
In competitive two-agent environments, deep reinforcement learning (RL) methods based on the \emph{Double Oracle (DO)} algorithm, such as \emph{Policy Space Response Oracles (PSRO)…