1 paper
Christian Bender, Nguyen Tran Thuan
We present a random measure approach for modeling exploration, i.e., the execution of measure-valued controls, in continuous-time reinforcement learning (RL) with controlled diffus…