1 paper
Li Jiang, Sijie Cheng, Jielin Qiu +3
The prevalent use of benchmarks in current offline reinforcement learning (RL) research has led to a neglect of the imbalance of real-world dataset distributions in the development…