1 paper
Haibin Zhou, Tong Wei, Zichuan Lin +6
We study the adaption of Soft Actor-Critic (SAC), which is considered as a state-of-the-art reinforcement learning (RL) algorithm, from continuous action space to discrete action s…