From the 1 of 1 linked paper with an AI index.
1 paper
Qi Feng, Gu Wang
The paper proposes a continuous policy‑value iteration method that uses Langevin‑type dynamics to update both the value function and the optimal control for stochastic control prob…