13 citations · 28 across the 10 of their papers we have counts for
1 paper · 1 filter
Stephen McAleer, JB Lanier, Kevin Wang +3
In competitive two-agent environments, deep reinforcement learning (RL) methods based on the \emph{Double Oracle (DO)} algorithm, such as \emph{Policy Space Response Oracles (PSRO)…