4 papers
Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories
Leo Benac, Abhishek Sharma, Sonali Parbhoo +1
We consider the problem of estimating the transition dynamics from near-optimal expert trajectories in the context of offline model-based reinforcement learning. We develop a…
Decision-Focused Model-based Reinforcement Learning for Reward Transfer
Abhishek Sharma, Sonali Parbhoo, Omer Gottesman +1
Model-based reinforcement learning (MBRL) provides a way to learn a transition model of the environment, which can then be used to plan personalized policies for different patient…
Decision-Point Guided Safe Policy Improvement
Abhishek Sharma, Leo Benac, Sonali Parbhoo +1
Within batch reinforcement learning, safe policy improvement (SPI) seeks to ensure that the learnt policy performs at least as well as the behavior policy that generated the datase…
Tree-Based Leakage Inspection and Control in Concept Bottleneck Models
Angelos Ragkousis, Sonali Parbhoo
As AI models grow larger, the demand for accountability and interpretability has become increasingly critical for understanding their decision-making processes. Concept Bottleneck…