2 papers
stat.ME2026
Conformalized Robust Principal Component Analysis
Liangliang Yuan, Lei Wang, Quan Kong +1
Robust principal component analysis (RPCA) is a widely used technique for recovering low-rank structure from matrices with missing entries and sparse, possibly large-magnitude corr…
cs.LG2025
SAMG: Offline-to-Online Reinforcement Learning via State-Action-Conditional Offline Model Guidance
Liyu Zhang, Haochi Wu, Xu Wan +3
Offline-to-online (O2O) reinforcement learning (RL) pre-trains models on offline data and refines policies through online fine-tuning. However, existing O2O RL algorithms typically…