1 paper
Germano Gabbianelli, Gergely Neu, Nneka Okolo +1
Offline Reinforcement Learning (RL) aims to learn a near-optimal policy from a fixed dataset of transitions collected by another policy. This problem has attracted a lot of attenti…