1 paper
Tatsuhiro Shimizu, Kazuki Kawamura, Takanori Muroi +4
We study the novel problem of future off-policy evaluation (F-OPE) and learning (F-OPL) for estimating and optimizing the future value of policies in non-stationary environments, w…