Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
CRAX: Fast Safe Reinforcement Learning Benchmarking
Tristan Tomilin, Mourad Boustani, Mickey Beurskens +1
Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While benchmarks have been central to progr…
cs.LG2026
Perception-Based Beliefs for POMDPs with Visual Observations
Miriam Schäfers, Merlijn Krale, Thiago D. Simão +2
Partially observable Markov decision processes (POMDPs) are a principled planning model for sequential decision-making under uncertainty. Yet, real-world problems with high-dimensi…