Showing cs.LGShow all
2 papers · 1 filter
cs.LG2019
Incrementally Learning Functions of the Return
Brendan Bennett, Wesley Chung, Muhammad Zaheer +1
Temporal difference methods enable efficient estimation of value functions in reinforcement learning in an incremental fashion, and are of broader interest because they correspond…
cs.LG2019
Planning with Expectation Models
Yi Wan, Zaheer Abbas, Adam White +2
Distribution and sample models are two popular model choices in model-based reinforcement learning (MBRL). However, learning these models can be intractable, particularly when the…