4 citations · 12 across the 32 of their papers we have counts for
1 paper · 1 filter
Fan Zhang, Baoru Huang, Xin Zhang
Offline reinforcement learning aims to learn an agent from pre-collected datasets, avoiding unsafe and inefficient real-time interaction. However, inevitable access to out-ofdistri…