From the 1 of 7 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Quantile Geometry Regularization for Distributional Reinforcement Learning
Zhaofan Zhang, Minghao Yang, Rufeng Chen +2
Quantile-based distributional reinforcement learning methods learn return distributions through sampled quantile regression, but their bootstrapped target quantiles may induce dist…
cs.LG2026
Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning
Rufeng Chen, Zhaofan Zhang, Zhejiang Yang +2
Offline safe reinforcement learning often requires policies to adapt at deployment time to safety budgets that vary across episodes or change within a single episode. While diffusi…