18 citations · 39 across the 15 of their papers we have counts for
4 papers · 1 filter
SAT-MARL: Specification Aware Training in Multi-Agent Reinforcement Learning
Fabian Ritz, Thomy Phan, Robert Müller +8
A characteristic of reinforcement learning is the ability to develop unforeseen strategies when solving problems. While such strategies sometimes yield superior performance, they m…
Uncertainty-Based Out-of-Distribution Classification in Deep Reinforcement Learning
Andreas Sedlmeier, Thomas Gabor, Thomy Phan +2
Robustness to out-of-distribution (OOD) data is an important goal in building reliable machine learning systems. Especially in autonomous systems, wrong predictions for OOD inputs…
Uncertainty-Based Out-of-Distribution Detection in Deep Reinforcement Learning
Andreas Sedlmeier, Thomas Gabor, Thomy Phan +2
We consider the problem of detecting out-of-distribution (OOD) samples in deep reinforcement learning. In a value based reinforcement learning setting, we propose to use uncertaint…
QoS-Aware Multi-Armed Bandits
Lenz Belzner, Thomas Gabor
Motivated by runtime verification of QoS requirements in self-adaptive and self-organizing systems that are able to reconfigure their structure and behavior in response to runtime…