2 papers
cs.AI2026
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
Haibo Tong, Feifei Zhao, Linghao Feng +18
Rapidly evolving AI exhibits increasingly strong autonomy and goal-directed capabilities, accompanied by derivative systemic risks that are more unpredictable, difficult to control…
cs.LG2025
A Variance-Reduced Cubic-Regularized Newton for Policy Optimization
Cheng Sun, Zhen Zhang, Shaofu Yang
In this paper, we study a second-order approach to policy optimization in reinforcement learning. Existing second-order methods often suffer from suboptimal sample complexity or re…