1 paper
Jaehyun Park, Junyeop Kwon, Dabeen Lee
We study model-based reinforcement learning with non-linear function approximation where the transition function of the underlying Markov decision process (MDP) is given by a multi…