1 paper
Enric Ribera Borrell, Lorenz Richter, Christof Schütte
We extend the standard reinforcement learning framework to random time horizons. While the classical setting typically assumes finite and deterministic or infinite runtimes of traj…