activity
20232026
most citedMulti-Agent Risks from Advanced AI

10 citations · 13 across the 14 of their papers we have counts for

collaborators
Showing cs.AIShow all

7 papers · 1 filter

cs.AI2026

Debate is efficient with your time

Jonah Brown-Cohen, Geoffrey Irving, Simon C. Marshall +3

AI safety via debate uses two competing models to help a human judge verify complex computational tasks. Previous work has established what problems debate can solve in principle,…

cs.AI2025

Avoiding Obfuscation with Prover-Estimator Debate

Jonah Brown-Cohen, Geoffrey Irving, Georgios Piliouras +3

Training powerful AI systems to exhibit desired behaviors hinges on the ability to provide accurate human supervision on increasingly complex tasks. A promising approach to this pr…

cs.AI2025

Plasticity as the Mirror of Empowerment

David Abel, Michael Bowling, André Barreto +13

Agents are minimally entities that are influenced by their past observations and act to influence future observations. This latter capacity is captured by empowerment, which has se…

cs.AI2025

Agency Is Frame-Dependent

David Abel, André Barreto, Michael Bowling +13

Agency is a system's capacity to steer outcomes toward a goal, and is a central topic of study across biology, philosophy, cognitive science, and artificial intelligence. Determini…

cs.AI20241 cited

A theory of appropriateness with applications to generative artificial intelligence

Joel Z. Leibo, Alexander Sasha Vezhnevets, Manfred Diaz +11

What is appropriateness? Humans navigate a multi-scale mosaic of interlocking notions of what is appropriate for different situations. We act one way with our friends, another with…

cs.AI2024

Neural Population Learning beyond Symmetric Zero-sum Games

Siqi Liu, Luke Marris, Marc Lanctot +3

We study computationally efficient methods for finding equilibria in n-player general-sum games, specifically ones that afford complex visuomotor skills. We show how existing metho…