6 citations · 9 across the 2 of their papers we have counts for
4 papers
Pessimistic Iterative Planning with RNNs for Robust POMDPs
Maris F. L. Galesloot, Marnix Suilen, Thiago D. Simão +4
Robust POMDPs extend classical POMDPs to incorporate model uncertainty using so-called uncertainty sets on the transition and observation functions, effectively defining ranges of…
Maintenance Strategies for Sewer Pipes with Multi-State Degradation and Deep Reinforcement Learning
Lisandro A. Jimenez-Roa, Thiago D. Simão, Zaharah Bukhsh +4
Large-scale infrastructure systems are crucial for societal welfare, and their effective management requires strategic forecasting and intervention methods that account for various…
Safe Reinforcement Learning From Pixels Using a Stochastic Latent Representation
Yannick Hogewind, Thiago D. Simao, Tal Kachman +1
We address the problem of safe reinforcement learning from pixel observations. Inherent challenges in such settings are (1) a trade-off between reward optimization and adhering to…
Safe Policy Improvement with an Estimated Baseline Policy
Thiago D. Simão, Romain Laroche, Rémi Tachet des Combes
Previous work has shown the unreliability of existing algorithms in the batch Reinforcement Learning setting, and proposed the theoretically-grounded Safe Policy Improvement with B…