1 paper · 1 filter
Nan Lu, Ethan Lee, James M. Robins +2
We study offline inference for the optimal value in reinforcement learning. Two new nuisances are derived as fixed points of a self-induced Bellman equation, in which we approximat…