6 citations · 8 across the 4 of their papers we have counts for
1 paper · 1 filter
Conor F. Hayes, Mathieu Reymond, Diederik M. Roijers +2
In many risk-aware and multi-objective reinforcement learning settings, the utility of the user is derived from the single execution of a policy. In these settings, making decision…