3 papers
cs.CE2025
Fitting Reinforcement Learning Model to Behavioral Data under Bandits
Hao Zhu, Jasper Hoffmann, Baohe Zhang +1
We consider the problem of fitting a reinforcement learning (RL) model to some given behavioral data under a multi-armed bandit environment. These models have received much attenti…
cs.LG2024
Revisiting Safe Exploration in Safe Reinforcement learning
David Eckel, Baohe Zhang, Joschka Bödecker
Safe reinforcement learning (SafeRL) extends standard reinforcement learning with the idea of safety, where safety is typically defined through the constraint of the expected cost…
cs.LG2024
Constrained Reinforcement Learning for Safe Heat Pump Control
Baohe Zhang, Lilli Frison, Thomas Brox +1
Constrained Reinforcement Learning (RL) has emerged as a significant research area within RL, where integrating constraints with rewards is crucial for enhancing safety and perform…