3 citations · 3 across the 1 of their papers we have counts for
1 paper
Hongyao Tang, Jianye Hao, Guangyong Chen +4
Value functions are crucial for model-free Reinforcement Learning (RL) to obtain a policy implicitly or guide the policy updates. Value estimation heavily depends on the stochastic…