1 citations · 3 across the 5 of their papers we have counts for
5 papers
A theory of appropriateness with applications to generative artificial intelligence
Joel Z. Leibo, Alexander Sasha Vezhnevets, Manfred Diaz +11
What is appropriateness? Humans navigate a multi-scale mosaic of interlocking notions of what is appropriate for different situations. We act one way with our friends, another with…
Approximating the Core via Iterative Coalition Sampling
Ian Gemp, Marc Lanctot, Luke Marris +11
The core is a central solution concept in cooperative game theory, defined as the set of feasible allocations or payments such that no subset of agents has incentive to break away…
Doing the right thing for the right reason: Evaluating artificial moral cognition by probing cost insensitivity
Yiran Mao, Madeline G. Reinecke, Markus Kunesch +4
Is it possible to evaluate the moral cognition of complex artificial agents? In this work, we take a look at one aspect of morality: `doing the right thing for the right reasons.'…
Heterogeneous Social Value Orientation Leads to Meaningful Diversity in Sequential Social Dilemmas
Udari Madhushani, Kevin R. McKee, John P. Agapiou +6
In social psychology, Social Value Orientation (SVO) describes an individual's propensity to allocate resources between themself and others. In reinforcement learning, SVO has been…
Diversity Through Exclusion (DTE): Niche Identification for Reinforcement Learning through Value-Decomposition
Peter Sunehag, Alexander Sasha Vezhnevets, Edgar Duéñez-Guzmán +2
Many environments contain numerous available niches of variable value, each associated with a different local optimum in the space of behaviors (policy space). In such situations i…