1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.GT2022★ 1 cited
Self-Play PSRO: Toward Optimal Populations in Two-Player Zero-Sum Games
Stephen McAleer, JB Lanier, Kevin Wang +3
In competitive two-agent environments, deep reinforcement learning (RL) methods based on the \emph{Double Oracle (DO)} algorithm, such as \emph{Policy Space Response Oracles (PSRO)…
cs.LG2022
Feasible Adversarial Robust Reinforcement Learning for Underspecified Environments
JB Lanier, Stephen McAleer, Pierre Baldi +1
Robust reinforcement learning (RL) considers the problem of learning policies that perform well in the worst case among a set of possible environment parameter values. In real-worl…
cs.LG2021
Target Entropy Annealing for Discrete Soft Actor-Critic
Yaosheng Xu, Dailin Hu, Litian Liang +3
Soft Actor-Critic (SAC) is considered the state-of-the-art algorithm in continuous action space settings. It uses the maximum entropy framework for efficiency and stability, and ap…