3 citations · 3 across the 2 of their papers we have counts for
1 paper · 1 filter
Hao Liang, Zhi-Quan Luo
We study the regret guarantee for risk-sensitive reinforcement learning (RSRL) via distributional reinforcement learning (DRL) methods. In particular, we consider finite episodic M…