7 citations · 16 across the 13 of their papers we have counts for
3 papers · 1 filter
Maintenance Strategies for Sewer Pipes with Multi-State Degradation and Deep Reinforcement Learning
Lisandro A. Jimenez-Roa, Thiago D. Simão, Zaharah Bukhsh +4
Large-scale infrastructure systems are crucial for societal welfare, and their effective management requires strategic forecasting and intervention methods that account for various…
More for Less: Safe Policy Improvement With Stronger Performance Guarantees
Patrick Wienhöft, Marnix Suilen, Thiago D. Simão +3
In an offline reinforcement learning setting, the safe policy improvement (SPI) problem aims to improve the performance of a behavior policy according to which sample data has been…
Efficient Sensitivity Analysis for Parametric Robust Markov Chains
Thom Badings, Sebastian Junges, Ahmadreza Marandi +2
We provide a novel method for sensitivity analysis of parametric robust Markov chains. These models incorporate parameters and sets of probability distributions to alleviate the of…