6 citations · 8 across the 4 of their papers we have counts for
1 paper · 1 filter
Conor F. Hayes, Mathieu Reymond, Diederik M. Roijers +2
In many risk-aware and multi-objective reinforcement learning settings, the utility of the user is derived from a single execution of a policy. In these settings, making decisions…