From the 1 of 1 linked paper with an AI index.
1 paper
Zeyu Bian, Ying Zhou, Yifan Cui
The paper introduces a method for offline reinforcement learning when the true actions are unobserved, using next-state information to estimate policy values and providing a robust…