1 paper
Omer Gottesman, Joseph Futoma, Yao Liu +4
Off-policy evaluation in reinforcement learning offers the chance of using observational data to improve future outcomes in domains such as healthcare and education, but safe deplo…