4 citations · 4 across the 1 of their papers we have counts for
1 paper
Jongmin Lee, Wonseok Jeon, Byung-Jun Lee +2
We consider the offline reinforcement learning (RL) setting where the agent aims to optimize the policy solely from the data without further environment interactions. In offline RL…