2 papers
cs.LG2026
Residuals-based Offline Reinforcement Learning
Qing Zhu, Xian Yu
Offline reinforcement learning (RL) has received increasing attention for learning policies from previously collected data without interaction with the real environment, which is p…
cs.LG2025
Reward Redistribution via Gaussian Process Likelihood Estimation
Minheng Xiao, Xian Yu
In many practical reinforcement learning tasks, feedback is only provided at the end of a long horizon, leading to sparse and delayed rewards. Existing reward redistribution method…