1 paper
Hanyang Zhao, Wenpin Tang, David D. Yao
We study reinforcement learning (RL) in the setting of continuous time and space, for an infinite horizon with a discounted objective and the underlying dynamics driven by a stocha…