4 papers
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model
Laixi Shi, Gen Li, Yuting Wei +3
This paper investigates model robustness in reinforcement learning (RL) to reduce the sim-to-real gap in practice. We adopt the framework of distributionally robust Markov decision…
Fast Computation of Optimal Transport via Entropy-Regularized Extragradient Methods
Gen Li, Yanxi Chen, Yu Huang +3
Efficient computation of the optimal transport distance between two distributions serves as an algorithm subroutine that empowers various applications. This paper develops a scalab…
Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning
Gen Li, Yuling Yan, Yuxin Chen +1
This paper studies reward-agnostic exploration in reinforcement learning (RL) -- a scenario where the learner is unware of the reward functions during the exploration stage -- and…
High-probability sample complexities for policy evaluation with linear function approximation
Gen Li, Weichen Wu, Yuejie Chi +3
This paper is concerned with the problem of policy evaluation with linear function approximation in discounted infinite horizon Markov decision processes. We investigate the sample…