156 citations · 156 across the 1 of their papers we have counts for
1 paper
Chunlin Chen, Daoyi Dong, Han-Xiong Li +2
The balance between exploration and exploitation is a key problem for reinforcement learning methods, especially for Q-learning. In this paper, a fidelity-based probabilistic Q-lea…