1 paper
Tyler Sam, Yudong Chen, Christina Lee Yu
Many reinforcement learning (RL) algorithms are too costly to use in practice due to the large sizes S,A of the problem's state and action space. To resolve this issue, we study…