18 citations · 33 across the 11 of their papers we have counts for
4 papers · 1 filter
Logic-based AI for Interpretable Board Game Winner Prediction with Tsetlin Machine
Charul Giri, Ole-Christoffer Granmo, Herke van Hoof +1
Hex is a turn-based two-player connection game with a high branching factor, making the game arbitrarily complex with increasing board sizes. As such, top-performing algorithms for…
Reliably Re-Acting to Partner's Actions with the Social Intrinsic Motivation of Transfer Empowerment
Tessa van der Heiden, Herke van Hoof, Efstratios Gavves +1
We consider multi-agent reinforcement learning (MARL) for cooperative communication and coordination tasks. MARL agents can be brittle because they can overfit their training partn…
An Autonomous Free Airspace En-route Controller using Deep Reinforcement Learning Techniques
Joris Mollinga, Herke van Hoof
Air traffic control is becoming a more and more complex task due to the increasing number of aircraft. Current air traffic control methods are not suitable for managing this increa…
Addressing Function Approximation Error in Actor-Critic Methods
Scott Fujimoto, Herke van Hoof, David Meger
In value-based reinforcement learning methods such as deep Q-learning, function approximation errors are known to lead to overestimated value estimates and suboptimal policies. We…