1 paper
Zilong Deng, Simon Khan, Shaofeng Zou
In this work, we study the sample complexity problem of risk-sensitive Reinforcement Learning (RL) with a generative model, where we aim to maximize the Conditional Value at Risk (…