2 papers
cs.LG2024
Offline Reinforcement Learning: Role of State Aggregation and Trajectory Data
Zeyu Jia, Alexander Rakhlin, Ayush Sekhari +1
We revisit the problem of offline reinforcement learning with value function realizability but without Bellman completeness. Previous work by Xie and Jiang (2021) and Foster et al.…
cs.LG2023
When is Agnostic Reinforcement Learning Statistically Tractable?
Zeyu Jia, Gene Li, Alexander Rakhlin +2
We study the problem of agnostic PAC reinforcement learning (RL): given a policy class , how many rounds of interaction with an unknown MDP (with a potentially large state and a…