2 papers
cs.LG2025
Model-Based Offline Reinforcement Learning with Adversarial Data Augmentation
Hongye Cao, Fan Feng, Jing Huo +4
Model-based offline Reinforcement Learning (RL) constructs environment models from offline datasets to perform conservative policy optimization. Existing approaches focus on learni…
cs.LG2024
A Variance Minimization Approach to Temporal-Difference Learning
Xingguo Chen, Yu Gong, Shangdong Yang +1
Fast-converging algorithms are a contemporary requirement in reinforcement learning. In the context of linear function approximation, the magnitude of the smallest eigenvalue of th…