4 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2021★ 4 cited
A Spectral Approach to Off-Policy Evaluation for POMDPs
Yash Nair, Nan Jiang
We consider off-policy evaluation (OPE) in Partially Observable Markov Decision Processes, where the evaluation policy depends only on observable variables but the behavior policy…
cs.LG2020★ 1 cited
PAC Bounds for Imitation and Model-based Batch Learning of Contextual Markov Decision Processes
Yash Nair, Finale Doshi-Velez
We consider the problem of batch multi-task reinforcement learning with observed context descriptors, motivated by its application to personalized medical treatment. In particular,…