activity
20242026
collaborators

12 papers

cs.GT2026

Paradoxes of Game Theoretic Equilibria and Price of Anarchy

Georgios Piliouras, Ian Gemp, Siqi Liu +1

The paper shows that static equilibrium concepts like Nash and correlated equilibria can hide unstable dynamics in multi‑agent learning, leading to unbounded or chaotic inefficienc…

cs.AI2026

Debate is efficient with your time

Jonah Brown-Cohen, Geoffrey Irving, Simon C. Marshall +3

AI safety via debate uses two competing models to help a human judge verify complex computational tasks. Previous work has established what problems debate can solve in principle,…

cs.LG2025

Cautious Optimism: A Meta-Algorithm for Near-Constant Regret in General Games

Ashkan Soleymani, Georgios Piliouras, Gabriele Farina

We introduce Cautious Optimism, a framework for substantially faster regularized learning in general games. Cautious Optimism, as a variant of Optimism, adaptively controls the lea…

cs.AI2025

Plasticity as the Mirror of Empowerment

David Abel, Michael Bowling, André Barreto +13

Agents are minimally entities that are influenced by their past observations and act to influence future observations. This latter capacity is captured by empowerment, which has se…

cs.AI2025

Avoiding Obfuscation with Prover-Estimator Debate

Jonah Brown-Cohen, Geoffrey Irving, Georgios Piliouras +3

Training powerful AI systems to exhibit desired behaviors hinges on the ability to provide accurate human supervision on increasingly complex tasks. A promising approach to this pr…

cs.GT2025

Convex Markov Games: A New Frontier for Multi-Agent Reinforcement Learning

Ian Gemp, Andreas Haupt, Luke Marris +2

Behavioral diversity, expert imitation, fairness, safety goals and others give rise to preferences in sequential decision making domains that do not decompose additively across tim…