4 citations · 4 across the 2 of their papers we have counts for
1 paper · 1 filter
Yingdong Lu, Mark S. Squillante, Chai Wah Wu
We consider a new form of reinforcement learning (RL) that is based on opportunities to directly learn the optimal control policy and a general Markov decision process (MDP) framew…