1 paper · 1 filter
Keiichiro Takahashi, Taisuke Kobayashi, Tomoya Yamanokuchi +1
This paper investigates a novel nonlinear update rule based on temporal difference (TD) errors in reinforcement learning (RL). The update rule in the standard RL states that the TD…