activity
20162020
most citedEmotion in Reinforcement Learning Agents and Robots: A Survey

159 citations · 183 across the 5 of their papers we have counts for

collaborators

12 papers

cs.AI20202 cited

The Second Type of Uncertainty in Monte Carlo Tree Search

Thomas M Moerland, Joost Broekens, Aske Plaat +1

Monte Carlo Tree Search (MCTS) efficiently balances exploration and exploitation in tree search based on count-derived uncertainty. However, these local visit counts ignore a secon…

cs.AI20206 cited

Think Too Fast Nor Too Slow: The Computational Trade-off Between Planning And Reinforcement Learning

Thomas M. Moerland, Anna Deichler, Simone Baldi +2

Planning and reinforcement learning are two key approaches to sequential decision making. Multi-step approximate real-time dynamic programming, a recently successful algorithm clas…

cs.MA2020

Automated Configuration of Negotiation Strategies

Bram M. Renting, Holger H. Hoos, Catholijn M. Jonker

Bidding and acceptance strategies have a substantial impact on the outcome of negotiations in scenarios with linear additive and nonlinear utility functions. Over the years, it has…

cs.MA2018

Towards Agent-based Models of Rumours in Organizations: A Social Practice Theory Approach

Amir Ebrahimi Fard, Rijk Mercuur, Virginia Dignum +2

Rumour is a collective emergent phenomenon with a potential for provoking a crisis. Modelling approaches have been deployed since five decades ago; however, the focus was mostly on…

cs.MA2018

Modelling Agents Endowed with Social Practices: Static Aspects

Rijk Mercuur, Virginia Dignum, Catholijn M. Jonker

To understand societal phenomena through simulation, we need computational variants of socio-cognitive theories. Social Practice Theory has provided a unique understanding of socia…

cs.LG201714 cited

Efficient exploration with Double Uncertain Value Networks

Thomas M. Moerland, Joost Broekens, Catholijn M. Jonker

This paper studies directed exploration for reinforcement learning agents by tracking uncertainty about the value of each available action. We identify two sources of uncertainty t…