1 paper
Samin Yeasar Arnob, Scott Fujimoto, Doina Precup
In this paper, we investigate the use of small datasets in the context of offline reinforcement learning (RL). While many common offline RL benchmarks employ datasets with over a m…