2 papers
cs.LG2024
Diffusion-DICE: In-Sample Diffusion Guidance for Offline Reinforcement Learning
Liyuan Mao, Haoran Xu, Xianyuan Zhan +2
One important property of DIstribution Correction Estimation (DICE) methods is that the solution is the optimal stationary distribution ratio between the optimized and data collect…
cs.LG2024
Offline Reinforcement Learning with Imbalanced Datasets
Li Jiang, Sijie Cheng, Jielin Qiu +3
The prevalent use of benchmarks in current offline reinforcement learning (RL) research has led to a neglect of the imbalance of real-world dataset distributions in the development…