convergence analysis 1entropy regularization 1newton-raphson method 1regularized policy iteration 1reinforcement learning 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.LG2026
Bridging the Gap between Newton-Raphson Method and Regularized Policy Iteration
Zeyang Li, Chuxiong Hu, Yunan Wang +4
The paper shows that regularized policy iteration in reinforcement learning is mathematically equivalent to applying the Newton‑Raphson method to a smoothed Bellman equation, provi…
cs.RO2026
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion
Guanchen Lu, Yajuan Dun, Yi Zhou +4
Scalable reinforcement learning has popularized high-throughput sampling architectures, which significantly compresses the training time for off-policy methods in robotic locomotio…
eess.SY2026
On the Optimization Landscape of Observer-based Dynamic Linear Quadratic Control
Jingliang Duan, Jie Li, Yinsong Ma +5
Understanding the optimization landscape of linear quadratic regulation (LQR) problems is fundamental to the design of efficient reinforcement learning solutions. Recent work has m…