16 citations · 17 across the 4 of their papers we have counts for
1 paper · 1 filter
Mridul Agarwal, Vaneet Aggarwal
Reinforcement learning typically assumes that the state update from the previous actions happens instantaneously, and thus can be used for making future decisions. However, this ma…