2 papers
cs.AI2025
Capturing Individual Human Preferences with Reward Features
André Barreto, Vincent Dumoulin, Yiran Mao +6
Reinforcement learning from human feedback usually models preferences using a reward function that does not distinguish between people. We argue that this is unlikely to be a good…
cs.GT2024
Approximating the Core via Iterative Coalition Sampling
Ian Gemp, Marc Lanctot, Luke Marris +11
The core is a central solution concept in cooperative game theory, defined as the set of feasible allocations or payments such that no subset of agents has incentive to break away…