6 citations · 8 across the 6 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2024★ 6 cited
Maintenance Strategies for Sewer Pipes with Multi-State Degradation and Deep Reinforcement Learning
Lisandro A. Jimenez-Roa, Thiago D. Simão, Zaharah Bukhsh +4
Large-scale infrastructure systems are crucial for societal welfare, and their effective management requires strategic forecasting and intervention methods that account for various…
cs.LG2023
More for Less: Safe Policy Improvement With Stronger Performance Guarantees
Patrick Wienhöft, Marnix Suilen, Thiago D. Simão +3
In an offline reinforcement learning setting, the safe policy improvement (SPI) problem aims to improve the performance of a behavior policy according to which sample data has been…