4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 4 cited
Is Vanilla Policy Gradient Overlooked? Analyzing Deep Reinforcement Learning for Hanabi
Bram Grooten, Jelle Wemmenhove, Maurice Poot +1
In pursuit of enhanced multi-agent collaboration, we analyze several on-policy deep reinforcement learning algorithms in the recently published Hanabi benchmark. Our research sugge…
eess.SY2020
On the Role of Models in Learning Control: Actor-Critic Iterative Learning Control
Maurice Poot, Jim Portegies, Tom Oomen
Learning from data of past tasks can substantially improve the accuracy of mechatronic systems. Often, for fast and safe learning a model of the system is required. The aim of this…