3 citations · 7 across the 4 of their papers we have counts for
4 papers
Percentile Criterion Optimization in Offline Reinforcement Learning
Elita A. Lobo, Cyrus Cousins, Yair Zick +1
In reinforcement learning, robust policies for high-stakes decision-making problems with limited data are usually computed by optimizing the \emph{percentile criterion}. The percen…
Data Poisoning Attacks on Off-Policy Policy Evaluation Methods
Elita Lobo, Harvineet Singh, Marek Petrik +2
Off-policy Evaluation (OPE) methods are a crucial tool for evaluating policies in high-stakes domains such as healthcare, where exploration is often infeasible, unethical, or expen…
Axiomatic Aggregations of Abductive Explanations
Gagan Biradar, Yacine Izza, Elita Lobo +2
The recent criticisms of the robustness of post hoc model approximation explanation methods (like LIME and SHAP) have led to the rise of model-precise abductive explanations. For e…
Soft Options Critic
Elita Lobo, Scott Jordan
The option-critic architecture (Bacon, Harb, and Precup 2017) and several variants have successfully demonstrated the use of the options framework proposed by Sutton et al (Sutton,…