2 papers
cs.LG2026
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning
Yuan Zhuang, Yuexin Bian, Sihong He +7
Scaling critic capacity is a promising direction for improving off-policy reinforcement learning (RL). However, recent work shows that larger critics are prone to overfitting and i…
cs.LG2025
CUQDS: Conformal Uncertainty Quantification under Distribution Shift for Trajectory Prediction
Huiqun Huang, Sihong He, Fei Miao
Trajectory prediction models that can infer both finite future trajectories and their associated uncertainties of the target vehicles in an online setting (e.g., real-world applica…