48 citations · 48 across the 1 of their papers we have counts for
1 paper
Gaon An, Seungyong Moon, Jang-Hyun Kim +1
Offline reinforcement learning (offline RL), which aims to find an optimal policy from a previously collected static dataset, bears algorithmic difficulties due to function approxi…