350 citations · 670 across the 13 of their papers we have counts for
1 paper · 1 filter
Stephen McAleer, JB Lanier, Kevin Wang +3
In competitive two-agent environments, deep reinforcement learning (RL) methods based on the \emph{Double Oracle (DO)} algorithm, such as \emph{Policy Space Response Oracles (PSRO)…