1 paper · 1 filter
Christian Bender, Nguyen Tran Thuan
We present a random measure approach for modeling exploration, i.e., the execution of measure-valued controls, in continuous-time reinforcement learning (RL) with controlled diffus…