11 citations · 12 across the 4 of their papers we have counts for
3 papers · 1 filter
Anytime PSRO for Two-Player Zero-Sum Games
Stephen McAleer, Kevin Wang, John Lanier +4
Policy space response oracles (PSRO) is a multi-agent reinforcement learning algorithm that has achieved state-of-the-art performance in very large two-player zero-sum games. PSRO…
Improving Social Welfare While Preserving Autonomy via a Pareto Mediator
Stephen McAleer, John Lanier, Michael Dennis +2
Machine learning algorithms often make decisions on behalf of agents with varied and sometimes conflicting interests. In domains where agents can choose to take their own action or…
Pipeline PSRO: A Scalable Approach for Finding Approximate Nash Equilibria in Large Games
Stephen McAleer, John Lanier, Roy Fox +1
Finding approximate Nash equilibria in zero-sum imperfect-information games is challenging when the number of information states is large. Policy Space Response Oracles (PSRO) is a…