Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Robust Probabilistic Shielding for Safe Offline Reinforcement Learning
Maris F. L. Galesloot, Thomas Rhemrev, Nils Jansen
In offline reinforcement learning (RL), we learn policies from fixed datasets without environment interaction. The major challenges are to provide guarantees on the (1) performance…
cs.LG2026
Perception-Based Beliefs for POMDPs with Visual Observations
Miriam Schäfers, Merlijn Krale, Thiago D. Simão +2
Partially observable Markov decision processes (POMDPs) are a principled planning model for sequential decision-making under uncertainty. Yet, real-world problems with high-dimensi…
cs.LG2024
Maintenance Strategies for Sewer Pipes with Multi-State Degradation and Deep Reinforcement Learning
Lisandro A. Jimenez-Roa, Thiago D. Simão, Zaharah Bukhsh +4
Large-scale infrastructure systems are crucial for societal welfare, and their effective management requires strategic forecasting and intervention methods that account for various…