1 paper
Matthew Landers, Taylor W. Killian, Hugo Barnes +2
Offline reinforcement learning in high-dimensional, discrete action spaces is challenging due to the exponential scaling of the joint action space with the number of sub-actions an…