1 paper
Aayam Shrestha, Stefan Lee, Prasad Tadepalli +1
We study an approach to offline reinforcement learning (RL) based on optimally solving finitely-represented MDPs derived from a static dataset of experience. This approach can be a…