1 paper
Yoshihiro Okawa, Tomotake Sasaki, Hidenao Iwane
In reinforcement learning (RL) algorithms, exploratory control inputs are used during learning to acquire knowledge for decision making and control, while the true dynamics of a co…