1 citations · 1 across the 1 of their papers we have counts for
1 paper
Jinyang Jiang, Jiaqiao Hu, Yijie Peng
Classical reinforcement learning (RL) aims to optimize the expected cumulative rewards. In this work, we consider the RL setting where the goal is to optimize the quantile of the c…