2 citations · 2 across the 1 of their papers we have counts for
1 paper
Allen Nie, Yannis Flet-Berliac, Deon R. Jordan +2
Offline reinforcement learning (RL) can be used to improve future performance by leveraging historical data. There exist many different algorithms for offline RL, and it is well re…