273 citations · 593 across the 26 of their papers we have counts for
1 paper · 1 filter
Xinyan Yan, Krzysztof Choromanski, Byron Boots +1
Policy evaluation or value function or Q-function approximation is a key procedure in reinforcement learning (RL). It is a necessary component of policy iteration and can be used f…