2 papers
cs.LG2024
Safe Time-Varying Optimization based on Gaussian Processes with Spatio-Temporal Kernel
Jialin Li, Marta Zagorowska, Giulia De Pasquale +2
Ensuring safety is a key aspect in sequential decision making problems, such as robotics or process control. The complexity of the underlying systems often makes finding the optima…
cs.LG2024
DTR-Bench: An in silico Environment and Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime
Zhiyao Luo, Mingcheng Zhu, Fenglin Liu +4
Reinforcement learning (RL) has garnered increasing recognition for its potential to optimise dynamic treatment regimes (DTRs) in personalised medicine, particularly for drug dosag…