3 papers
cs.LG2025
Bellman Error Centering
Xingguo Chen, Yu Gong, Shangdong Yang +1
This paper revisits the recently proposed reward centering algorithms including simple reward centering (SRC) and value-based reward centering (VRC), and points out that SRC is ind…
cs.LG2022
Keeping Minimal Experience to Achieve Efficient Interpretable Policy Distillation
Xiao Liu, Shuyang Liu, Wenbin Li +2
Although deep reinforcement learning has become a universal solution for complex control tasks, its real-world applicability is still limited because lacking security guarantees fo…
cs.LG2022
Online Attentive Kernel-Based Temporal Difference Learning
Guang Yang, Xingguo Chen, Shangdong Yang +3
With rising uncertainty in the real world, online Reinforcement Learning (RL) has been receiving increasing attention due to its fast learning capability and improving data efficie…