2 papers
cs.LG2025
Reward Redistribution via Gaussian Process Likelihood Estimation
Minheng Xiao, Xian Yu
In many practical reinforcement learning tasks, feedback is only provided at the end of a long horizon, leading to sparse and delayed rewards. Existing reward redistribution method…
cs.LG2025
Policy Gradient Methods for Risk-Sensitive Distributional Reinforcement Learning with Provable Convergence
Minheng Xiao, Xian Yu, Lei Ying
Risk-sensitive reinforcement learning (RL) is crucial for maintaining reliable performance in high-stakes applications. While traditional RL methods aim to learn a point estimate o…