7 citations · 21 across the 6 of their papers we have counts for
1 paper · 1 filter
Conor F. Hayes, Mathieu Reymond, Diederik M. Roijers +2
In many risk-aware and multi-objective reinforcement learning settings, the utility of the user is derived from a single execution of a policy. In these settings, making decisions…