2 papers
cs.LG2025
On the Convergence of Single-Timescale Actor-Critic
Navdeep Kumar, Priyank Agrawal, Giorgia Ramponi +2
We analyze the global convergence of the single-timescale actor-critic (AC) algorithm for the infinite-horizon discounted Markov Decision Processes (MDPs) with finite state spaces.…
cs.AI2025
Dual Formulation for Non-Rectangular Lp Robust Markov Decision Processes
Navdeep Kumar, Adarsh Gupta, Maxence Mohamed Elfatihi +3
We study robust Markov decision processes (RMDPs) with non-rectangular uncertainty sets, which capture interdependencies across states unlike traditional rectangular models. While…