most citedDoing the right thing for the right reason: Evaluating artificial moral cognition by probing cost insensitivity

1 citations · 3 across the 5 of their papers we have counts for

collaborators

5 papers

cs.AI20241 cited

A theory of appropriateness with applications to generative artificial intelligence

Joel Z. Leibo, Alexander Sasha Vezhnevets, Manfred Diaz +11

What is appropriateness? Humans navigate a multi-scale mosaic of interlocking notions of what is appropriate for different situations. We act one way with our friends, another with…

cs.GT2024

Approximating the Core via Iterative Coalition Sampling

Ian Gemp, Marc Lanctot, Luke Marris +11

The core is a central solution concept in cooperative game theory, defined as the set of feasible allocations or payments such that no subset of agents has incentive to break away…

cs.AI20231 cited

Doing the right thing for the right reason: Evaluating artificial moral cognition by probing cost insensitivity

Yiran Mao, Madeline G. Reinecke, Markus Kunesch +4

Is it possible to evaluate the moral cognition of complex artificial agents? In this work, we take a look at one aspect of morality: `doing the right thing for the right reasons.'…

cs.MA20231 cited

Heterogeneous Social Value Orientation Leads to Meaningful Diversity in Sequential Social Dilemmas

Udari Madhushani, Kevin R. McKee, John P. Agapiou +6

In social psychology, Social Value Orientation (SVO) describes an individual's propensity to allocate resources between themself and others. In reinforcement learning, SVO has been…

cs.AI2023

Diversity Through Exclusion (DTE): Niche Identification for Reinforcement Learning through Value-Decomposition

Peter Sunehag, Alexander Sasha Vezhnevets, Edgar Duéñez-Guzmán +2

Many environments contain numerous available niches of variable value, each associated with a different local optimum in the space of behaviors (policy space). In such situations i…