10 citations · 10 across the 11 of their papers we have counts for
3 papers · 1 filter
When (and How) to Trust the Expert: Diagnosing Query-Time Expert-Guided Reinforcement Learning
Yann Berthelot, Philippe Preux, Riad Akrour
Many continuous-control problems ship with a competent but suboptimal controller (a tuned PID, a hand-designed gait). A growing family of methods uses such controllers as queryable…
PB: Preference Space Exploration via Population-Based Methods in Preference-Based Reinforcement Learning
Brahim Driss, Alex Davey, Riad Akrour
Preference-based reinforcement learning (PbRL) has emerged as a promising approach for learning behaviors from human feedback without predefined reward functions. However, current…
Interpretable and Editable Programmatic Tree Policies for Reinforcement Learning
Hector Kohler, Quentin Delfosse, Riad Akrour +2
Deep reinforcement learning agents are prone to goal misalignments. The black-box nature of their policies hinders the detection and correction of such misalignments, and the trust…