1 paper match
Qi Feng, Gu Wang
The paper proposes a continuous policy‑value iteration method that uses Langevin‑type dynamics to update both the value function and the optimal control for stochastic control prob…