2 citations · 2 across the 1 of their papers we have counts for
1 paper
Qiang He, Xinwen Hou
Offline reinforcement learning (RL), also known as batch RL, aims to optimize policy from a large pre-recorded dataset without interaction with the environment. This setting offers…