1 paper
Maris F. L. Galesloot, Thomas Rhemrev, Nils Jansen
In offline reinforcement learning (RL), we learn policies from fixed datasets without environment interaction. The major challenges are to provide guarantees on the (1) performance…