5 citations · 5 across the 2 of their papers we have counts for
3 papers
stat.ML2024
Sequential Harmful Shift Detection Without Labels
Salim I. Amoukou, Tom Bewley, Saumitra Mishra +3
We introduce a novel approach for detecting distribution shifts that negatively impact the performance of machine learning models in continuous production environments, which requi…
cs.AI2023
Learning Interpretable Models of Aircraft Handling Behaviour by Reinforcement Learning from Human Feedback
Tom Bewley, Jonathan Lawry, Arthur Richards
We propose a method to capture the handling abilities of fast jet pilots in a software model via reinforcement learning (RL) from human preference feedback. We use pairwise prefere…
cs.LG2021★ 5 cited
Interpretable Preference-based Reinforcement Learning with Tree-Structured Reward Functions
Tom Bewley, Freddy Lecue
The potential of reinforcement learning (RL) to deliver aligned and performant agents is partially bottlenecked by the reward engineering problem. One alternative to heuristic tria…