activity
20162026
most citedEmotion in Reinforcement Learning Agents and Robots: A Survey

159 citations · 230 across the 22 of their papers we have counts for

collaborators
Showing 2018Show all

6 papers · 1 filter

cs.MA2018

Towards Agent-based Models of Rumours in Organizations: A Social Practice Theory Approach

Amir Ebrahimi Fard, Rijk Mercuur, Virginia Dignum +2

Rumour is a collective emergent phenomenon with a potential for provoking a crisis. Modelling approaches have been deployed since five decades ago; however, the focus was mostly on…

cs.MA2018

Modelling Agents Endowed with Social Practices: Static Aspects

Rijk Mercuur, Virginia Dignum, Catholijn M. Jonker

To understand societal phenomena through simulation, we need computational variants of socio-cognitive theories. Social Practice Theory has provided a unique understanding of socia…

cs.LG2018

The Potential of the Return Distribution for Exploration in RL

Thomas M. Moerland, Joost Broekens, Catholijn M. Jonker

This paper studies the potential of the return distribution for exploration in deterministic reinforcement learning (RL) environments. We study network losses and propagation mecha…

stat.ML2018

Monte Carlo Tree Search for Asymmetric Trees

Thomas M. Moerland, Joost Broekens, Aske Plaat +1

We present an extension of Monte Carlo Tree Search (MCTS) that strongly increases its efficiency for trees with asymmetry and/or loops. Asymmetric termination of search trees intro…

cs.MA2018

Volunteers in the Smart City: Comparison of Contribution Strategies on Human-Centered Measures

Stefano Bennati, Ivana Dusparic, Rhythima Shinde +1

Several smart city services rely on users contribution, e.g., data, which can be costly for the users in terms of privacy. High costs lead to reduced user participation, which unde…

cs.LG2018

Ordered Preference Elicitation Strategies for Supporting Multi-Objective Decision Making

Luisa M Zintgraf, Diederik M Roijers, Sjoerd Linders +2

In multi-objective decision planning and learning, much attention is paid to producing optimal solution sets that contain an optimal policy for every possible user preference profi…