6 citations · 21 across the 8 of their papers we have counts for
5 papers · 1 filter
An interpretable semi-supervised classifier using two different strategies for amended self-labeling
Isel Grau, Dipankar Sengupta, Maria M. Garcia Lorenzo +1
In the context of some machine learning applications, obtaining data instances is a relatively easy process but labeling them could become quite expensive or tedious. Such scenario…
Sample-Efficient Model-Free Reinforcement Learning with Off-Policy Critics
Denis Steckelmacher, Hélène Plisnier, Diederik M. Roijers +1
Value-based reinforcement-learning algorithms provide state-of-the-art results in model-free discrete-action settings, and tend to outperform actor-critic algorithms. We argue that…
Dynamic Weights in Multi-Objective Deep Reinforcement Learning
Axel Abels, Diederik M. Roijers, Tom Lenaerts +2
Many real-world decision problems are characterized by multiple conflicting objectives which must be balanced based on their relative importance. In the dynamic weights setting the…
Directed Policy Gradient for Safe Reinforcement Learning with Human Advice
Hélène Plisnier, Denis Steckelmacher, Tim Brys +2
Many currently deployed Reinforcement Learning agents work in an environment shared with humans, be them co-workers, users or clients. It is desirable that these agents adjust to p…
Ordered Preference Elicitation Strategies for Supporting Multi-Objective Decision Making
Luisa M Zintgraf, Diederik M Roijers, Sjoerd Linders +2
In multi-objective decision planning and learning, much attention is paid to producing optimal solution sets that contain an optimal policy for every possible user preference profi…