1 paper
Daiqi Gao, Hsin-Yu Lai, Predrag Klasnja +1
We consider reinforcement learning (RL) for a class of problems with bagged decision times. A bag contains a finite sequence of consecutive decision times. The transition dynamics…