4 papers
Simulating Eutopia: Revisiting Long-term Fairness with Outcomes, Performativity, and Dynamics
Vedant Palit, Udvas Das, Brahim Driss +1
As AI-driven Decision Makers (ADMs) influence our socioeconomic reality, their roles in both enhancing efficiency and amplifying the social biases have drawn attention. In this pap…
Performative Policy Gradient: Optimality in Performative Reinforcement Learning
Debabrota Basu, Udvas Das, Brahim Driss +1
Post-deployment machine learning algorithms often influence the environments they act in, and thus shift the underlying dynamics that the standard reinforcement learning (RL) metho…
StaQ: a Finite Memory Approach to Discrete Action Policy Mirror Descent
Alena Shilova, Alex Davey, Brahim Driss +1
In Reinforcement Learning (RL), regularization with a Kullback-Leibler divergence that penalizes large deviations between successive policies has emerged as a popular tool both in…
PB: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning
Brahim Driss, Alex Davey, Riad Akrour
Preference-based reinforcement learning (PbRL) has emerged as a promising approach for learning behaviors from human feedback without predefined reward functions. However, current…