paper

On the grid-sampling limit SDE

arXiv:2410.07778

Abstract

In our recent work [3] we introduced the grid-sampling SDE as a proxy for modeling exploration in continuous-time reinforcement learning. In this note, we provide further motivation for the use of this SDE and discuss its wellposedness in the presence of jumps.

This note provides supplementary materials to arXiv:2409.17200 in a self-contained way

On the grid-sampling limit SDE · wovepaper