1 paper · 1 filter
Brendan Bennett, Wesley Chung, Muhammad Zaheer +1
Temporal difference methods enable efficient estimation of value functions in reinforcement learning in an incremental fashion, and are of broader interest because they correspond…