3 papers
math.OC2026
Optimal local interventions in the two-dimensional Abelian sandpile model
Maike C. de Jongh, Richard J. Boucherie, M. N. M. van Lieshout
The Abelian sandpile model serves as a canonical example of self-organized criticality. This critical behavior manifests itself through large cascading events triggered by small pe…
stat.ML2025
Convergence of off-policy TD(0) with linear function approximation for reversible Markov chains
Maik Overmars, Jasper Goseling, Richard Boucherie
We study the convergence of off-policy TD(0) with linear function approximation when used to approximate the expected discounted reward in a Markov chain. It is well known that the…
math.OC2025
Controlling the low-temperature Ising model using spatiotemporal Markov decision theory
M. C. de Jongh, Richard J. Boucherie, M. N. M. van Lieshout
We introduce the spatiotemporal Markov decision process (STMDP), a special type of Markov decision process that models sequential decision-making problems which are not only charac…