6 citations · 13 across the 10 of their papers we have counts for
4 papers · 1 filter
Safe Reinforcement Learning with Instantaneous Constraints: The Role of Aggressive Exploration
Honghao Wei, Xin Liu, Lei Ying
This paper studies safe Reinforcement Learning (safe RL) with linear function approximation and under hard instantaneous constraints where unsafe actions must be avoided at each st…
Model-Free, Regret-Optimal Best Policy Identification in Online CMDPs
Zihan Zhou, Honghao Wei, Lei Ying
This paper considers the best policy identification (BPI) problem in online Constrained Markov Decision Processes (CMDPs). We are interested in algorithms that are model-free, have…
Sample Efficient Reinforcement Learning in Mixed Systems through Augmented Samples and Its Applications to Queueing Networks
Honghao Wei, Xin Liu, Weina Wang +1
This paper considers a class of reinforcement learning problems, which involve systems with two types of states: stochastic and pseudo-stochastic. In such systems, stochastic state…
Provably Efficient Model-Free Algorithms for Non-stationary CMDPs
Honghao Wei, Arnob Ghosh, Ness Shroff +2
We study model-free reinforcement learning (RL) algorithms in episodic non-stationary constrained Markov Decision Processes (CMDPs), in which an agent aims to maximize the expected…