3 citations · 7 across the 3 of their papers we have counts for
3 papers
math.OC2021★ 3 cited
Faster Algorithm and Sharper Analysis for Constrained Markov Decision Process
Tianjiao Li, Ziwei Guan, Shaofeng Zou +3
The problem of constrained Markov decision process (CMDP) is investigated, where an agent aims to maximize the expected accumulated discounted reward subject to multiple constraint…
cs.LG2020★ 3 cited
When Will Generative Adversarial Imitation Learning Algorithms Attain Global Convergence
Ziwei Guan, Tengyu Xu, Yingbin Liang
Generative adversarial imitation learning (GAIL) is a popular inverse reinforcement learning approach for jointly optimizing policy and reward from expert trajectories. A primary q…
cs.LG2020★ 1 cited
Robust Stochastic Bandit Algorithms under Probabilistic Unbounded Adversarial Attack
Ziwei Guan, Kaiyi Ji, Donald J Bucci +4
The multi-armed bandit formalism has been extensively studied under various attack models, in which an adversary can modify the reward revealed to the player. Previous studies focu…