1 paper
Sibo Zhang, Rui Jing, Liangfu Lv +2
Reinforcement learning (RL) has achieved notable performance in high-dimensional sequential decision-making tasks, yet remains limited by low sample efficiency, sensitivity to nois…