3 papers
cs.LG2025
On the Effect of Regularization in Policy Mirror Descent
Jan Felix Kleuker, Aske Plaat, Thomas Moerland
Policy Mirror Descent (PMD) has emerged as a unifying framework in reinforcement learning (RL) by linking policy gradient methods with a first-order optimization method known as mi…
cs.LG2025
Chargax: A JAX Accelerated EV Charging Simulator
Koen Ponse, Jan Felix Kleuker, Aske Plaat +1
Deep Reinforcement Learning can play a key role in addressing sustainable energy challenges. For instance, many grid systems are heavily congested, highlighting the urgent need to…
cs.LG2024
Reinforcement Learning for Sustainable Energy: A Survey
Koen Ponse, Felix Kleuker, Márton Fejér +3
The transition to sustainable energy is a key challenge of our time, requiring modifications in the entire pipeline of energy production, storage, transmission, and consumption. At…