1 paper
Matteo Gallici, Mattie Fellows, Benjamin Ellis +4
Q-learning played a foundational role in the field reinforcement learning (RL). However, TD algorithms with off-policy data, such as Q-learning, or nonlinear function approximation…