2 papers
cs.LG2025
CAWR: Corruption-Averse Advantage-Weighted Regression for Robust Policy Optimization
Ranting Hu
Offline reinforcement learning (offline RL) algorithms often require additional constraints or penalty terms to address distribution shift issues, such as adding implicit or explic…
cs.DS2024
Computing the Center of Uncertain Points on Cactus Graphs
Ran Hu, Divy H. Kanani, Jingru Zhang
In this paper, we consider the (weighted) one-center problem of uncertain points on a cactus graph. Given are a cactus graph and a set of uncertain points. Each uncertain p…