activity
20162024
most citedModel-Based Reinforcement Learning with Value-Targeted Regression

69 citations · 523 across the 40 of their papers we have counts for

collaborators
Showing 2023 · cs.LGShow all

9 papers · 2 filters

cs.LG2023★ 1 cited

Sample Complexity of Neural Policy Mirror Descent for Policy Optimization on Low-Dimensional Manifolds

Zhenghao Xu, Xiang Ji, Minshuo Chen +2

Policy gradient methods equipped with deep neural networks have achieved great success in solving high-dimensional reinforcement learning (RL) problems. However, current analyses c…

cs.LG2023★ 1 cited

Deep Reinforcement Learning for Efficient and Fair Allocation of Health Care Resources

Yikuan Li, Chengsheng Mao, Kaixuan Huang +4

Scarcity of health care resources could result in the unavoidable consequence of rationing. For example, ventilators are often limited in supply, especially during public health em…

cs.LG2023

Actions Speak What You Want: Provably Sample-Efficient Reinforcement Learning of the Quantal Stackelberg Equilibrium from Strategic Feedbacks

Siyu Chen, Mengdi Wang, Zhuoran Yang

We study reinforcement learning (RL) for learning a Quantal Stackelberg Equilibrium (QSE) in an episodic Markov game with a leader-follower structure. In specific, at the outset of…

cs.LG2023★ 1 cited

Effective Minkowski Dimension of Deep Nonparametric Regression: Function Approximation and Statistical Theories

Zixuan Zhang, Minshuo Chen, Mengdi Wang +2

Existing theories on deep nonparametric regression have shown that when the input data lie on a low-dimensional manifold, deep neural networks can adapt to the intrinsic data struc…

cs.LG2023

Provably Efficient Representation Learning with Tractable Planning in Low-Rank POMDP

Jiacheng Guo, Zihao Li, Huazheng Wang +3

In this paper, we study representation learning in partially observable Markov Decision Processes (POMDPs), where the agent learns a decoder function that maps a series of high-dim…

cs.LG2023

Efficient Reinforcement Learning with Impaired Observability: Learning to Act with Delayed and Missing State Observations

Minshuo Chen, Jie Meng, Yu Bai +3

In real-world reinforcement learning (RL) systems, various forms of {\it impaired observability} can complicate matters. These situations arise when an agent is unable to observe t…