2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.AI2019★ 2 cited
Distributed Policy Iteration for Scalable Approximation of Cooperative Multi-Agent Policies
Thomy Phan, Kyrill Schmid, Lenz Belzner +3
Decision making in multi-agent systems (MAS) is a great challenge due to enormous state and joint action spaces as well as uncertainty, making centralized control generally infeasi…
cs.LG2019
Uncertainty-Based Out-of-Distribution Detection in Deep Reinforcement Learning
Andreas Sedlmeier, Thomas Gabor, Thomy Phan +2
We consider the problem of detecting out-of-distribution (OOD) samples in deep reinforcement learning. In a value based reinforcement learning setting, we propose to use uncertaint…