3 citations · 4 across the 4 of their papers we have counts for
1 paper · 1 filter
Yuhao Ding, Ming Jin, Javad Lavaei
We study risk-sensitive reinforcement learning (RL) based on an entropic risk measure in episodic non-stationary Markov decision processes (MDPs). Both the reward functions and the…