1 paper
Thomas G. Dietterich, George Trimponias, Zhitang Chen
Exogenous state variables and rewards can slow down reinforcement learning by injecting uncontrolled variation into the reward signal. We formalize exogenous state variables and re…