1 paper
Jiaxuan Hou, Lifeng Wei
We develop a continuous-time entropy-regularized reinforcement learning framework under model uncertainty. By applying Sion's minimax theorem, we transform the intractable robust c…