39 citations · 59 across the 21 of their papers we have counts for
Showing 2024Show all
2 papers · 1 filter
cs.AI2024
On the Limitations of Markovian Rewards to Express Multi-Objective, Risk-Sensitive, and Modal Tasks
Joar Skalse, Alessandro Abate
In this paper, we study the expressivity of scalar, Markovian reward functions in Reinforcement Learning (RL), and identify several limitations to what they can express. Specifical…
math.OC2024
Policy Evaluation in Distributional LQR (Extended Version)
Zifan Wang, Yulong Gao, Siyi Wang +3
Distributional reinforcement learning (DRL) enhances the understanding of the effects of the randomness in the environment by letting agents learn the distribution of a random retu…