1 paper · 1 filter
Shihao Li, Jiachen Li, Jiamin Xu +3
We study how trajectory value depends on the learning algorithm in policy-gradient control. Using Trajectory Shapley in an uncertain LQR, we find a negative correlation between Per…