1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Defu Cao, Angela Zhou
Offline reinforcement learning enables evaluation and optimization of sequential decisions from historical data, when it is not possible to deploy new policies online due to safety…