From the 1 of 1 linked paper with an AI index.
1 paper
Zeyang Li, Chuxiong Hu, Yunan Wang +4
The paper shows that regularized policy iteration in reinforcement learning is mathematically equivalent to applying the Newton‑Raphson method to a smoothed Bellman equation, provi…