5 citations · 9 across the 4 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2024
Policy Gradient for Robust Markov Decision Processes
Qiuhao Wang, Shaohang Xu, Chin Pang Ho +1
We develop a generic policy gradient method with the global optimality guarantee for robust Markov Decision Processes (MDPs). While policy gradient methods are widely used for solv…
cs.LG2022★ 2 cited
Policy Gradient in Robust MDPs with Global Convergence Guarantee
Qiuhao Wang, Chin Pang Ho, Marek Petrik
Robust Markov decision processes (RMDPs) provide a promising framework for computing reliable policies in the face of model errors. Many successful reinforcement learning algorithm…
cs.LG2021★ 2 cited
Optimization Induced Equilibrium Networks
Xingyu Xie, Qiuhao Wang, Zenan Ling +4
Implicit equilibrium models, i.e., deep neural networks (DNNs) defined by implicit equations, have been becoming more and more attractive recently. In this paper, we investigate an…