4 citations · 4 across the 1 of their papers we have counts for
1 paper
Nicholas E. Corrado, Yuxiao Qu, John U. Balis +2
In offline reinforcement learning (RL), an RL agent learns to solve a task using only a fixed dataset of previously collected data. While offline RL has been successful in learning…