1 paper · 1 filter
Andrew Bennett, Nathan Kallus, Miruna Oprescu +2
We study the evaluation of a policy under best- and worst-case perturbations to a Markov decision process (MDP), using transition observations from the original MDP, whether they a…